基础设施与安全主管
Head of Infrastructure & Security
Albert 通过构建理解化学家工作方式的技术,帮助化学家改变世界。作为化学家,我们知道当软件不符合你的工作方式时是什么感觉——被困在由从未在实验台工作过的人构建的系统中。这就是我们创建 Albert OS 的原因,这是首个面向材料科学的原生人工智能操作系统,旨在消除从好奇心到发现之间的所有障碍。
我们正在寻找一位基础设施与安全负责人,负责在 Albert 建立和主导全球基础设施、可靠性和安全职能。这是一个构建项目的角色,而不是继承项目的角色。你将处于云架构、站点可靠性工程和企业安全的交汇点,直接向工程副总裁汇报。
如果你认为正确的基础设施是一种竞争优势,我们希望听到你的声音。
你将做的事情
- 所有工程团队的云架构:通用的 IaC 工具链、统一监控、CI/CD 标准化和内部开发者平台。
- Albert 的端到端安全态势:补丁、SAST/DAST、WAF、SIEM、EDR、IDS/IPS、数据防泄漏和跨云与身份的零信任。你还将负责第三方安全合作、企业审计和问卷回复,以及公司范围内的安全意识。
- 与财务部门合作的云成本策略:单位指标、团队级预算可见性、供应商合同,以及面向管理层的基础设施效率仪表盘。
- 按照随日而动的模式建立全球 SRE 职能:现有的印度团队、美国存在,以及计划中的以德国为中心的欧洲团队。
- 机器学习/AI 基础设施平台,数据科学团队在其上构建。
- 与合作伙伴 Azure 环境的互操作性:身份联合、私有连接和自备密钥。
你在第一年将交付的内容
- 一个 SLO 框架、错误预算模型和服务分级分类,为工程和管理层提供关于可靠性的共同语言。
- 一个运行中的事件管理程序:定义了升级路径,安排了值班轮换,文档化了灾难恢复和业务连续性策略,并完成了第一次完整的事后分析。
- 一个拥有明确区域责任的 SRE 组织,并制定了欧洲扩展路线图。
- MLOps 基础设施由 AI/ML 团队移交,这样数据科学团队可以发布模型而不必管理后端基础设施。
- 我们首次企业安全审计顺利通过,已制定整改措施。
查看英文原文
Albert helps chemists change the world by building technology that understands how they work. As chemists ourselves, we know what it’s like when software doesn’t work the way you do — stuck with systems built by people who never spent a day at the bench. That’s why we created Albert OS, the first AI-native operating system for materials science, designed to remove every obstacle between curiosity and discovery.
We're looking for a Head of Infrastructure & Security to build and own the global infrastructure, reliability, and security function at Albert. This is a program-building role, not a program-inheriting one. You'll sit at the intersection of cloud architecture, site reliability engineering, and enterprise security, reporting directly to the VP of Engineering.
If you believe infrastructure done right is a competitive advantage, we want to hear from you.
What you'll do
- Cloud architecture across all engineering teams: common IaC tooling, unified monitoring, CI/CD standardization, and the internal developer platform.
- Albert's security posture end-to-end: patching, SAST/DAST, WAF, SIEM, EDR, IDS/IPS, data loss prevention, and Zero Trust across cloud and identity. You'll also own the third-party security partnership, enterprise audit and questionnaire response, and security awareness across the company.
- Cloud cost strategy in partnership with Finance: per-unit metrics, team-level budget visibility, vendor contracts, and leadership-facing dashboards on infrastructure efficiency.
- The global SRE function on a follow-the-sun model: an existing India team, US presence, and a planned Germany-anchored EU team.
- The ML/AI infrastructure platform that the data science team builds on.
- Interoperability with partner Azure environments: identity federation, private connectivity, and bring-your-own-key.
What you'll deliver in the first year
- An SLO framework, error budget model, and service tiering classification that give engineering and leadership a shared language for reliability trade-offs.
- A functioning incident program: escalation paths defined, on-call rotations in place, DR and business continuity strategy documented, and the first post-mortem run end-to-end.
- An SRE organization with clear regional ownership and a roadmap for the EU buildout.
- MLOps infrastructure ownership off the AI/ML team, so the data science team ships models without managing backend infrastructure.
- Our first enterprise security audit passed with an owned, defensible posture rather than a manually assembled response.
- A unified cloud architecture with measurable cost reduction against a defined per-unit baseline.
You will have
- 8+ years in cloud infrastructure, platform engineering, or site reliability engineering, with at least 3 years in a senior leadership role owning a team or function.
- Deep, current AWS expertise: ECS, multi-account organization design, IAM and KMS, CloudTrail-based audit evidence.
- Container orchestration at architectural depth, decision-making rather than operations. ECS and EKS are both in production here, and the consolidation question between them is live.
- Experience building SLO frameworks, error budgets, incident response, and DR programs from scratch, not inheriting them.
- Experience building or managing global, distributed engineering teams across time zones.
- Strong security fundamentals: cloud security architecture, IAM, Zero Trust, network segmentation, SAST/DAST, SIEM/EDR, and working knowledge of SOC 2, ISO 27001, NIST, or equivalent.
- A track record of cloud cost reduction through FinOps: unit economics, per-team budget accountability, and leadership-facing visibility.
- Startup or high-growth experience, and the ability to distinguish tech debt that matters from tech debt that can wait.
- Preferred: MLOps infrastructure (Kubeflow, Airflow, or equivalent), defense or government-adjacent compliance experience (NIST 800-171, CMMC, ITAR, FedRAMP), and relevant certifications (CISSP, CISM, CCSP).
Key competencies
Systems Thinking:Ability to visualize and design how AI agent systems connect, share data, and enable valuable customer outcomes at scale.
Grit:Self-starter, persistent problem solver, relentless in the pursuit of progress in ambiguous 0-to-1 product environments.
Customer Empathy:Deep understanding of how enterprise customersoperate; asks the right questions anddesignssolutions that stick.
Innovator:Passionate about applying emerging AI capabilities to practical, measurable business problems before the patterns are written.
Adaptability:Comfortable operating in dynamic, fast-moving customer and company environments where the approach evolves withnew information.
Leadership:Ability to drive agent skill initiatives end-to-end with minimal oversight, aligning cross-functional teams to shared outcomes.
Why Albert?
We havea huge impact. Albert is a growing team with a global reach.We’rebuilding toward a future where discovery feels like exploration, not administration. Where chemists can pursue their most ambitious ideas without friction.
We love our global team. Albert’s home-base is in the California Bay Area, but we have multiple offices and employees sprinkled around the globe: India,Germanyand Japan. In fact, over 50% of our employees work outside of California. An international remote culture is in our DNA.
We care about you. Albert works hard to create a positive environment for our employees, and we think your life outside of work is important too. We work hardand weplay hard.
We value diversity. Growing andmaintainingour inclusive and diverse team matters to us. We are committed to being a company where our employees feel comfortable bringing their authenticselvesto work and have the ability to be successful every day.
We’realways looking for humble, sharp, and creative folks to join the Albert team. If you think you might bea fit, please apply.
Originally posted on Himalayas