云基础设施总监
Director, Cloud Infrastructure
在 Sanity.io http://Sanity.io,我们正在打造下一代人工智能驱动的内容运营。我们的 AI 内容操作系统让团队可以按业务需求建模、创建和自动化内容,加速数字开发并提升内容运营效率。SKIMS、Figma、Riot Games、Anthropic、COMPLEX、Nordstrom 和 Morningbrew 等公司正在使用 Sanity 来驱动和自动化他们的内容运营。
Sanity 的基础设施正进入新阶段。下一阶段将明确所有权,强化平台基础,提升运营纪律性,并构建一个能支撑公司未来几阶段增长的架构。
我们正在寻找一位云基础设施总监来领导这项工作。
该职位将负责 Sanity 工程师每天所依赖的平台基础:云基础设施、Kubernetes、网络、路由、可观测性、CI/CD、部署路径、事件响应,以及跨产品团队实现生产环境责任的规范。
规模是真实的。Content Lake 每秒处理约 75,000 个请求,每分钟约 450 万个请求。Sanity 还在 CDN、边缘、网关、缓存、对象存储和 GCP 基础设施上运行关键路径。一些工作已经在进行中:将 Varnish 和 Mead 迁移到 Fastly,加强可观测性,完成开发者轮班制度的部署,并使生产就绪成为发布流程中的常态。
这是一个需要深入技术能力的领导职位。你应能与优秀的基础设施工程师进行技术探讨,做出艰难的架构决策,并围绕他们建立运营模式:平台集中负责哪些内容,SRE 通过团队实现哪些支持,以及产品团队如何自信地部署和运行服务。
你将负责以下工作:
- 制定 Sanity 下一阶段扩展的基础设施策略,制定清晰的路线图,涵盖云基础设施、可靠性、部署、可观测性、安全性、成本和开发者体验。
- 领导负责 Sanity 工程速度背后共享基础的团队:GCP 项目和集群、Kubernetes、网络、服务发现、路由、网关、CI/CD、基础设施即代码和可观测性。
- 明确划分平台和 SRE 的职责。平台应负责共享基础。SRE 应通过强大的工具、标准和事件支持,帮助产品团队良好地运行服务。
- 提升 Sanity 全体的可靠性标准
查看英文原文
At Sanity.io http://Sanity.io, we’re building the future of AI-powered Content Operations. Our AI Content Operating System gives teams the freedom to model, create, and automate content the way their business works, accelerating digital development and supercharging content operations efficiency. Companies like SKIMS, Figma, Riot Games, Anthropic, COMPLEX, Nordstrom, and Morningbrew are using Sanity to power and automate their content operations.
Sanity's infrastructure is entering a new stage. The next layer is clearer ownership, stronger platform foundations, better operational discipline, and an architecture that can carry the company through the next few stages of growth.
We're looking for a Director of Cloud Infrastructure to lead that work.
This person will own the platform foundations Sanity engineers build on every day: cloud infrastructure, Kubernetes, networking, routing, observability, CI/CD, deployment paths, incident response, and the standards that make production ownership work across product teams.
The scale is real. Content Lake alone handles around 75,000 requests a second, about 4.5m requests a minute. Sanity also runs critical paths across CDN, edge, gateway, caching, object storage, and GCP infrastructure. Some of the work is already in motion: moving Varnish and Mead onto Fastly, tightening observability, finishing our developer on-call rollout, and making production readiness a normal part of shipping.
This is a leadership role for someone who can still go deep technically. You should be able to spar with strong infrastructure engineers, make hard architectural calls, and build the operating model around them: what Platform owns centrally, what SRE enables through teams, and how product teams deploy and run their services with confidence.
WHAT YOU WILL DO
- Set the infrastructure strategy for Sanity's next stage of scale, with a clear roadmap across cloud infrastructure, reliability, deployment, observability, security, cost, and developer experience.
- Lead the teams responsible for the shared foundations behind Sanity's engineering velocity: GCP projects and clusters, Kubernetes, networking, service discovery, routing, gateways, CI/CD, infrastructure as code, and observability.
- Draw a clean line between Platform and SRE. Platform should own the shared foundations. SRE should help product teams run services well, with strong tooling, standards, and incident support.
- Raise the reliability bar across Sanity's production systems, including dashboards, alert severity, paging standards, service ownership, on-call readiness, and incident response.
- Make deployment boring in the best way: clear golden paths, production readiness checks, safe rollouts, useful automation, and fewer places engineers need to look before they can ship.
- Partner with Product, Engineering, Security, Support, Sales, and Customer Success on the infrastructure work that matters to customers: uptime, latency, scale, compliance, trust, and cost.
- Own cloud cost discipline without slowing the business down. You will need to understand where the money goes, make tradeoffs visible, and help teams build with cost in mind.
- Shape Sanity's longer-term architecture for multi-region scale, disaster recovery, data residency, and trust requirements like SOC 2, ISO 27001, PCI, HIPAA, or similar customer expectations.
- Hire, coach, and stretch infrastructure leaders and engineers. The team needs direction, high standards, and someone who can make strong technical people better.
ABOUT YOU
- You have led Infrastructure, Platform, SRE, Cloud, or Developer Platform teams in a scaling SaaS, cloud, infrastructure, API, data, or developer-tools company.
- You have operated systems with high request volume, multi-region production, strict uptime expectations, large cloud bills, customer-facing incidents, and trust requirements that made engineering quality visible.
- You have built or run production platforms with Kubernetes, GCP or AWS, Terraform or similar infrastructure as code, service discovery, networking, API gateways, CDNs, observability, CI/CD, and incident response.
- You are technically deep enough to debate architecture with senior infrastructure engineers and practical enough to make the call when the perfect answer is wasting time.
- You know how to split central platform ownership from product-team service ownership without creating a ticket queue that everyone resents.
- You improve on-call and reliability by building systems, standards, and feedback loops that make production healthier over time.
- You can turn messy infrastructure work into a strategy people can follow, then keep pushing until the work ships.
- You communicate clearly with executives, product leaders, engineers, and customer-facing teams. You can explain tradeoffs without sanding off the technical truth.
- You have managed managers or senior technical leads, and you know when to coach, when to set the bar, and when to get directly involved.
WHAT WE CAN OFFER:
- Real infrastructure scale and a clear mandate to change how it works.
- A senior seat in R&D, close to Product, Engineering, Security, and customer-facing teams.
- A highly skilled, inspiring, and supportive team.
- A positive, flexible, and trust-based work environment that supports long-term professional and personal growth.
- A global, culturally diverse group of colleagues and customers.
- Comprehensive health plans and perks.
- A healthy work-life balance that accommodates individual and family needs.
- Competitive stock options and location-based salary.
WHO WE ARE:
Sanity.io http://Sanity.io is a modern, flexible content operating system that replaces rigid legacy content management systems. One of our big differentiators is treating content as data so that it can be stored in a single source of truth, but seamlessly adapted and personalized for any channel without extra effort. Forward-thinking companies choose Sanity because they can create tailored content authoring experiences, customized workflows, and content models that reflect their business.
Sanity recently raised an $85m Series C led by GP Bullhound and is also backed by leading investors like ICONIQ Growth, Threshold Ventures, Heavybit and Shopify, as well as founders of companies like Vercel, WPEngine, Twitter, Mux, Netlify and Heroku. This funding round has put Sanity in a strong position for accelerated growth in the coming years.
You can only build a great company with a great culture. Sanity is a 200+ person company with highly committed and ambitious people. We are pioneers, we exist for our customers, we are hel ved, and we love type two fun! Read more about our values here! https://www.sanity.io/blog/our-sanity-values
Sanity.io pledges to be an organization that reflects the globally diverse audience that our product serves. We believe that in addition to hiring the best talent, a diversity of perspectives, ideas, and cultures leads to the creation of better products and services. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, or gender identity.
#LI-ST