解决方案架构师
Solutions Architect
ABOUT BASETEN
Baseten 为全球最具活力的 AI 公司提供关键推理支持,例如 Cursor、Notion、OpenEvidence、Abridge、Clay、Gamma 和 Writer。通过结合应用 AI 研究、灵活的基础设施和无缝的开发者工具,我们使处于 AI 前沿的公司能够将前沿模型投入生产。我们正在快速成长,并最近完成了 1.5 亿美元的 F 轮融资 https://www.baseten.co/blog/announcing-our-series-f/,由 Altimeter Capital、Conviction Partners 和 Spark Capital 领投。加入我们,帮助构建工程师们用来交付 AI 产品的平台。
THE ROLE
作为 Baseten 的解决方案架构师,你将与销售团队和客户紧密合作,将业务需求转化为技术方案,进行技术探索,并指导客户的可重复部署和价值验证。这个职位适合具有创业精神、面向客户的工程技术专业人士,他们希望深入了解现代公司如何大规模采用 AI,并享受在技术探索、方案设计、演示、部署规划和与销售及工程团队密切合作的现场客户实施工作。
RESPONSIBILITIES
- 与销售团队一起参与客户探索通话(通常是第二次通话,有时是大型账户的第一轮通话)。
- 主导演示和技术规划,以对齐成功标准、架构和部署方法。
- 负责基准测试和可重复部署,包括:
- 处理多种模式(LLM、嵌入、图像和视频生成、语音 AI 等)的标准部署模式和配置。
- 就 H100 与 B200 之间的权衡以及延迟优化与吞吐量优化的设置提供建议。
- 推动常见模型和用例的一致“手册”式部署。
- 成为 vLLM、SGLang 和 TRT-LLM 等不同运行时的高级用户,并了解它们之间常见的配置和权衡。
- 推动 POC 和项目执行,包括:
- 规划 POC 并保持利益相关者对时间表、交付成果和下一步计划的同步。
- 在 POC 中担任“牵头人”或项目经理。
- 在需要更深入或更复杂的工程技术工作时,引入 Forward Deployed Engineering (FDE) 的支持。
REQUIREMENTS
- AI/ML 背景,能够与技术利益相关者可信地讨论 AI/ML 相关话题。
- 强大的面向客户的沟通能力,包括能够进行结构化讨论。
查看英文原文
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products.
THE ROLE
As a Solutions Architect at Baseten you will partner closely with Sales and customers to translate business needs into technical solutions, run technical discovery, and guide repeatable deployments and proofs of value for customers. This role is a great fit for entrepreneurial, customer-facing technical professionals who want a front-row view into how modern companies adopt AI at scale, and who enjoy working across technical discovery, solution design, demos, deployment scoping, and hands-on customer implementations, in close partnership with Sales and Engineering.
RESPONSIBILITIES
- Partner with Sales on customer discovery calls (most often second calls, occasionally first calls for large accounts).
- Lead demos and technical scoping to align on success criteria, architecture, and deployment approach.
- Own benchmarking and repeatable deployments, including:
- Handling standard deployment patterns and configurations across many modalities – LLMs, embeddings, image and video generation, Voice AI, etc.
- Advising on tradeoffs like H100s vs B200s and latency-optimized vs throughput-optimized setups.
- Driving consistent “playbook” style deployments for common models and use cases.
- Become a power user of different runtimes such as vLLM, SGLang, and TRT-LLM and all the common configurations and tradeoffs between them
- Drive POC and project execution, including:
- Scoping POCs and keeping stakeholders aligned on timeline, deliverables, and next steps.
- Acting as the “ringleader” or project manager for POCs.
- Pulling in Forward Deployed Engineering (FDE) support when deeper or more complex technical work is needed.
REQUIREMENTS
- AI/ML background and the ability to credibly discuss AI/ML topics with technical stakeholders.
- Strong customer-facing communication skills, including the ability to run structured discovery and clarify ambiguous requirements.
- Technical depth to scope solutions, without needing to write production code.
- Ability to script and prototype as needed, including comfort “vibe coding” to move quickly in technical workflows.
NICE TO HAVE
- Experience running or supporting benchmarks for ML inference deployments.
- Familiarity with infrastructure tradeoffs relevant to inference performance and cost (for example GPU selection and latency versus throughput tuning).
- Experience serving as a cross-functional technical lead for customer POCs, including coordination across Sales and Engineering.
BENEFITS
- Competitive compensation, including meaningful equity
- (U.S. only) 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).