首席工程师, 云基础设施
Principal Engineer, Cloud Infrastructure
我们正在构建支撑医疗行业最具雄心的技术转型的云基础设施和交付平台,我们需要一位高级工程师,负责云基础设施,全面掌控所有基础设施。你将既设定技术方向,又亲自参与工作,同时带领一支小型的基础设施工程师团队。
这个职位关注基础设施和开发者赋能。随着我们的AI战略不断发展,我们看到一个世界,许多传统工程和数据角色之外的人也将为我们的技术生态系统做出贡献。我们认为可靠性、安全性和成本纪律来自于良好的平台设计和自动化,而工单队列和运行手册是未解决根本问题的标志。你将对现代、面向AI的基础设施组织应具备的形态有清晰的观点,具备将其变为现实的判断力和技术能力,并具备领导力来培养和提升你周围的贡献者。
#### 在这个职位中,你将:
- 全面负责我们云环境中的云基础设施,与团队一起亲自参与IaC、CI/CD和可观测性
- 负责SDLC的范围(开发环境、低级环境、发布机制、监控和响应),让工程师能够在几分钟而不是几天内进行发布
- 构建让AI代理能够观察、推理并作用于业务系统的基础架构,以及让非技术人士开发的AI辅助应用能够被部署和支持而不成为运营负担的框架
- 作为基础设施方面的负责人,与我们的SecOps团队一起参与安全运营,包括CSPM、威胁管理和补丁管理
- 以业务成果(运营效率、财务影响、临床结果)来阐述基础设施决策,并将进展和权衡作为你工作方式的自然结果向领导层展示
- 领导并发展一支小型高影响力团队,并作为统一的技术力量,与我们的SecOps、数据、IT系统和应用工程领导者平级协作
#### 如果你有以下背景,请联系我们:
- 有工程背景,拥有在至少两个主要云平台(GCP、AWS、Azure)上操作生产基础设施的深入、实际经验,以及现代基础设施即代码(Terraform或类似工具)的经验
- 有构建和运维现代CI/CD系统、开发环境和可观测性堆栈的经验。你所交付的内容比规模更重要
查看英文原文
We're building the cloud infrastructure and delivery platform that one of healthcare's most ambitious technology transformations runs on, and we need a Principal Engineer, Cloud Infrastructure to own the infrastructure behind it all. You'll both set the technical direction and do the work hands-on, while leading a small team of infrastructure engineers.
This role is about infrastructure and developer enablement. As our AI strategy evolves, we see a world where many outside of traditional engineering and data roles will contribute to our technology ecosystem. We believe reliability, security, and cost discipline come from good platform design and automation, and that ticket queues and runbooks are signs we haven't solved the underlying problem. You'll bring a clear point of view on what a modern, AI-native infrastructure org should look like, the judgment and technical skill to make it real, and the leadership instincts to build and develop the contributors around you.
#### In this role you will:
- Own cloud infrastructure end-to-end across our cloud environments, staying hands-on with IaC, CI/CD, and observability alongside the team
- Own the SDLC surface area (dev environments, lower envs, release mechanics, monitoring and response) so engineers can ship in minutes rather than days
- Build the substrate that lets AI agents observe, reason, and act on business systems with safe defaults, and the rails that let AI-assisted apps from people outside of tech be deployed and supported without becoming operational liabilities
- Participate in security operations alongside our SecOps team on CSPM, threat management, and patch management as the infrastructure-side owner of remediations
- Frame infrastructure decisions in terms of business outcomes (operational efficiency, financial impact, clinical results), and make progress and tradeoffs visible to leadership as a natural byproduct of how you work
- Lead and develop a small, high-impact team, and operate as a peer to our SecOps, Data, IT Systems, and App Engineering leaders as a unifying technical force across the org
#### You should get in touch if:
- You have an engineering background with deep, hands-on experience operating production infrastructure across at least two major clouds (GCP, AWS, Azure) and with modern infrastructure-as-code (Terraform or similar)
- You have built and operated modern CI/CD systems, dev environments, and observability stacks. What you've shipped matters more than the size of the fleet you've managed.
- You've built and operated infrastructure across different company sizes and business contexts. Cross-domain breadth is a real asset; experience in regulated domains (healthcare, financial services) is helpful but not required
- You hold a strong, well-reasoned point of view on platform philosophy, and can defend when to standardize, when to give teams room, and how to make the safe path the easy path
- You have shipped AI-native operational systems. That might be AI agents that triage alerts, draft RCAs, manage cost, or execute runbooks, or frameworks that let non-traditional contributors (junior engineers, analysts, AI agents, vibe-coders) ship safely.
- You're action-oriented on governance and security participation. You document what matters, automate what you can, and don't let process become a bottleneck
- You have built and developed teams. You think about composition intentionally, hire for gaps, and invest in people's growth
- You're a strong communicator: clear, direct, and concise with both technical and non-technical audiences
#### Success in this role looks like
- First 90 Days: Develop a clear point of view on the current state of the infrastructure, CI/CD, and operational tooling. Identify the highest-leverage opportunities. Ship a first step-function improvement or eliminate a fundamental operational limitation.
- 6 Months: AI-assisted operations are in place for alert triage and incident response. CI/CD and dev environment strategy are coherent across the org. The small team operates in a new operating model with substantially higher leverage per engineer.
- Long Term: Infrastructure is a quiet enabler. Engineers, AI agents, and non-tech contributors ship safely and quickly without the platform being the bottleneck. Cloud cost is managed as an engineering discipline. The team is growing skills, increasing capacity, and operating with autonomy.
_We're hiring across multiple teams at Clover and Counterpart Health— our recruiting team will assess your background and connect you with the best-fit opportunity. You'll learn more about the specific team and interview process as you move through the process._
* * *
**Benefits Overview**:
- **Financial Well-Being**: Our commitment to attracting and retaining top talent begins with a competitive base salary and equity opportunities. Additionally, we offer a performance-based bonus program, 401k matching, and regular compensation reviews to recognize and reward exceptional contributions.
- **Physical Well-Being**: We prioritize the health and well-being of our employees and their families by providing comprehensive medical, dental, and vision coverage. Your health matters to us, and we invest in ensuring you have access to quality healthcare.
- **Mental Well-Being**: We understand the importance of mental health in fostering productivity and maintaining work-life balance. To support this, we offer initiatives such as No-Meeting Fridays, monthly company holidays, access to mental health resources, and a generous flexible time-off policy. Additionally, we embrace a remote-first culture that supports collaboration and flexibility, allowing our team members to thrive from any location.
- **Professional Development**: Developing internal talent is a priority for Clover. We offer learning programs, mentorship, professional development funding, and regular performance feedback and reviews.
_Additional Perks:_
- Employee Stock Purchase Plan (ESPP) offering discounted equity opportunities
- Reimbursement for office setup expenses
- Monthly cell phone & internet stipend
- Remote-first culture, enabling collaboration with global teams
- Paid parental leave for all new parents
- And much more!
* * *
**About Clover:** We are reinventing health insurance by combining the power of data with human empathy to keep our members healthier. We believe the healthcare system is broken, so we've created custom software and analytics to empower our clinical staff to intervene and provide personalized care to the people who need it most.
We always put our members first, and our success as a team is measured by the quality of life of the people we serve. Those who work at Clover are passionate and mission-driven individuals with diverse areas of expertise, working together to solve the most complicated problem in the world: healthcare.
From Clover’s inception, Diversity & Inclusion have always been key to our success. We are an Equal Opportunity Employer and our employees are people with different strengths, experiences, perspectives, opinions, and backgrounds, who share a passion for improving people's lives. Diversity not only includes race and gender identity, but also age, disability status, veteran status, sexual orientation, religion and many other parts of one’s identity. All of our employee’s points of view are key to our success, and inclusion is everyone's responsibility.
* * *
#LI-Remote
_Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records._ _We are an [E-Verify](https://www.e-verify.gov/?utm_medium=search&utm_source=google&utm_campaign=everify2018&utm_content=bg_Branded_Everify_General_English_BMM_E_Verify&utm_keyword=everify) company._
* * *
Final pay is based on several factors including but not limited to internal equity, market data, and the applicant’s education, work experience, certifications, etc.
A reasonable estimate of the base salary range for this role is:
$195,000—$270,000 USD