远程工作雷达

基础设施与站点可靠性工程负责人

Lead Infrastructure & Site Reliability Engineering

开发工程职能支持限定地区(需当地身份)
公司Crum & Forster
薪资$105,400 - $198,100/年
工作地点United States
地域资格限定地区(需当地身份)
时区要求日间重叠约 9 小时,基本正常作息
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 United States 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

Crum & Forster 公司简介
Travel Insured International(TII),Crum & Forster 公司旗下,正在招聘一位高级基础设施与站点可靠性工程师。
Travel Insured International 是一家领先的旅行保险提供商,拥有超过 30 年的经营历史。作为我们特种业务单元的重要组成部分,隶属于意外与健康部门,TII 为每位个人提供旅行保障计划,帮助其安心出行。Travel Insured International 非常自豪地为各类消费者和代理合作伙伴提供产品。我们致力于为所有客户提供可靠保障、优质价值和全流程满意。
职位描述
这是一个具有强大领导影响力的技术主导岗位,专注于可靠性和平台。你将通过技术可信度和影响力来领导——制定标准、塑造方向、提升工程团队的可靠性能力,同时保持深入参与。作为 TII 的高级基础设施与站点可靠性工程师,你将负责运行时的可靠性、可观测性、云安全和平台工程,确保我们的全 Azure 环境安全、弹性且持续可用,支持平台和业务的运行。
这是一个构建与增强型岗位:你将提升我们在指标、日志和追踪方面的可观测性;建立规范的事件指挥和无责复盘实践;使用 Bicep 和 Terraform 编写基础设施代码;强化 Azure 平台的安全态势;并构建自助服务平台功能,使工程团队能够安全地快速推进。你将负责可靠性、安全性和平台的实战工作,同时与工程架构、工程领导和安全相关方合作。与基础设施副总裁(AVP)共同制定 Cloud、Observability、ITSM 及相关领域的未来愿景和路线图,使其成为支持 TII 增长目标和拓展新分销渠道的共享、可衡量的工程学科。
你将负责:
基础设施策略

  • 负责并演进 TII 的 Azure 基础设施和可靠性路线图,符合 Azure Well-Architected Framework。
  • 与基础设施副总裁(AVP)共同制定 Cloud、Observability、ITSM 及相关领域的未来愿景和路线图。
  • 定义可随业务增长和新渠道扩展的计算、网络和平台服务标准。
  • 推动云成本优化(FinOps)——平衡性能、弹性和成本。
查看英文原文

Crum & Forster Company Overview
Travel Insured International (TII), a Crum & Forster company, is hiring a Lead Infrastructure & Site Reliability Engineer.
Travel Insured International is a leading travel insurance provider with more than 30 years in business. As a key component of our Specialty Business Unit, within the Accident & Health division, TII provides travel protection plans to help each individual travel confidently. Travel Insured International is proud to offer products to consumers and to agency partners of all sizes. We're committed to providing dependable coverage, great value, and end-to-end satisfaction for all customers.
Job Description
This is a hands-on engineering role with strong leadership influence, focused on reliability and platform. You will lead through technical credibility and influence—setting standards, shaping direction, and elevating the reliability capability across engineering—while remaining deeply hands-on. As TII's Lead Infrastructure & Site Reliability Engineer, you will own the run-time reliability, observability, cloud security, and platform engineering that keep our all-Azure environment secure, resilient, and always-on—supporting the platform and business.
This is a build-and-enhance role: you will mature our observability across metrics, logs, and traces; establish disciplined incident command and blameless post-incident practices; codify infrastructure with Bicep and Terraform; harden the security posture of our Azure platform; and build the self-service platform capabilities that let engineering teams move fast safely. You will own reliability, security, and platform hands-on while partnering with engineering architecture, engineering leadership, and security stakeholders. With the AVP, Infrastructure, co-create the future-state vision and roadmap for Cloud, Observability, ITSM, and related domains making it a shared, measurable engineering discipline in support of TII's growth target and expansion into new distribution channels.
What you will do:
Infrastructure Strategy

  • Own and evolve TII's Azure infrastructure and reliability roadmap, aligned to the Azure Well-Architected Framework.
  • With the AVP, Infrastructure to co-create the future-state vision and roadmap for Cloud, Observability, ITSM, and related domains.
  • Define standards for compute, network, and platform services that scale with business growth and new channels.
  • Drive cloud cost optimization (FinOps)—balancing performance, resilience, and spend.
  • Partner with engineering architecture to ensure infrastructure enables design-time resilience and delivery velocity.

Site Reliability Engineering

  • Own run-time reliability across availability, performance, scalability, and capacity for TII's platform.
  • Mature and expand the SLO practice—defining SLIs, refining the 99.9% (and higher, where warranted) SLOs, and operating error budgets to balance reliability and delivery speed.
  • Lead capacity planning and performance engineering to support the platform's growth.
  • Drive operational readiness reviews for new services and major releases.

Observability

  • Own and mature the observability platform across the three pillars—metrics, logs, and traces—enhancing Grafana/Prometheus and Azure Application Insights.
  • Implement distributed tracing across the GraphQL/REST services to accelerate diagnosis and reduce time-to-detect and time-to-resolve.
  • Establish meaningful alerting and telemetry that reduce noise and surface real signals.
  • Build reliability dashboards that give teams and leadership clear visibility into service health.

Platform Engineering

  • Own the internal developer platform and self-service infrastructure capabilities that enable engineering teams to provision and operate safely.
  • Define golden paths / paved-road templates that make the reliable, secure, and compliant way the easy way.
  • Establish and champion Infrastructure as Code standards using Bicep and Terraform.
  • Improve developer experience and engineering enablement through automation and reusable platform services.

Operational Excellence & Automation

  • Drive automation across provisioning, configuration, deployment, and remediation to eliminate toil.
  • Establish operational runbooks, self-healing patterns, and proactive reliability practices.
  • Continuously improve deployment safety and rollback capability in partnership with CI/CD owners.

Cloud Security & Compliance

  • Own the engineering and operational security of the Azure cloud platform—including identity and access management, network security, configuration hardening, and secrets/key management.
  • Manage and improve the cloud security posture (e.g., Microsoft Defender for Cloud, Azure Policy), including continuous vulnerability management and remediation.
  • Implement security monitoring and alerting as part of the observability platform to detect and respond to threats.
  • Embed DevSecOps and secure-by-design practices into platform and IaC workflows, enforcing guardrails and policy-as-code within golden paths and self-service tooling.
  • Partner with the Security function on policy, governance, and compliance in a regulated insurance (PII) environment.

Incident & Problem Management

  • Own the major-incident process and incident command, matured on the on-call platform (Better Stack).
  • Lead blameless post-incident reviews and drive systemic problem management to prevent recurrence.
  • Improve on-call health, escalation paths, and mean-time-to-detect / mean-time-to-resolve.
  • Maintain and enhance the mature Business Continuity and Disaster Recovery capabilities, including RTO/RPO targets and periodic testing.

Leadership

  • Will co-lead a small team of infrastructure, cloud, & system engineers.
  • Uplift the reliability and platform capability across engineering—raising standards and building a reliability culture.
  • Mentor engineers, demonstrating the leadership behaviors that support growth into a formal infrastructure leadership role.
  • Establish standards, documentation, and ways of working that scale across teams.
  • Other duties as assigned

What YOU will bring to C&F:
· Excellent problem-solving and analytical skills with attention to detail

  • Strong analytical and problem-solving skills.
  • Excellent verbal and written communication skills, with the ability to explain technical and functional issues clearly to both technical and non-technical stakeholders.
  • Demonstrated leadership or mentoring of distributed/offshore teams (formal people-management experience a plus).
  • Excellent collaboration and influencing skills
  • Outcome & Metrics Orientation
  • Self-starter

Requirements:

  • Bachelor's degree in Computer Science, or related field—or equivalent experience.
  • 8+ years in infrastructure, site reliability, platform, or DevOps engineering, with recent hands-on delivery.
  • Deep, hands-on expertise operating production workloads on Microsoft Azure.
  • Proven experience with Infrastructure as Code using Bicep and/or Terraform.
  • Hands-on experience with observability tooling across metrics, logs, and traces (e.g., Grafana, Prometheus, Azure Application Insights).
  • Proven experience defining and operating SLIs/SLOs and error budgets.
  • Hands-on experience securing Azure cloud environments—identity/access, network security, posture management (e.g., Microsoft Defender for Cloud, Azure Policy), and vulnerability management.
  • Experience scaling Microsoft Fabric and Microsoft Purview.
  • Experience owning incident management, on-call, and blameless post-incident reviews (e.g., Better Stack, PagerDuty, or comparable).
  • Strong scripting/automation ability (e.g., PowerShell, Python, Bash) and CI/CD experience (Azure DevOps preferred).
  • Experience supporting business continuity and disaster recovery with defined RTO/RPO.
  • Experience building internal developer platforms, golden paths, and self-service infrastructure.
  • Cloud cost optimization / FinOps experience.
  • Exposure to AI-assisted operations (AIOps) and modern reliability automation.
  • Microsoft Azure certifications (e.g., Azure Solutions Architect Expert, Azure DevOps Engineer Expert, Azure Security Engineer Associate) preferred.
  • Experience in insurance, travel, fintech, or SaaS (regulated environments) preferred.
  • Prior leadership or mentoring of distributed/offshore teams desired. Formal people-leadership experience a plus

What C&F will bring to you
What C&F will bring to YOU:

  • Competitive compensation package
  • Generous 401K employer match
  • Employee Stock Purchase plan with employer matching
  • Generous Paid Time Off
  • Excellent benefits that go beyond health, dental & vision. Our programs are focused on your whole family’s wellness including your physical, mental and financial wellbeing
  • A core C&F tenant is owning your career development so we provide a wealth of ways for you to keep learning, including tuition reimbursement, industry related certifications and professional training to keep you progressing on your chosen path
  • A dynamic, ambitious, fun and exciting work environment
  • We believe you do well by doing good and want to encourage a spirit of social and community responsibility, matching donation program, volunteer opportunities, and an employee driven corporate giving program that lets you participate and support your community

At C&F you will BELONG
We value inclusivity and diversity. We are committed to equal employment opportunity and welcome everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. If you require a special accommodation, please let us know.
For California Residents Only: Information collected and processed as part of your career profile and any job applications you choose to submit are subject to our privacy notices and policies, visit for more information.
Crum & Forster is committed to ensuring a workplace free from discriminatory pay disparities and complying with applicable pay equity laws. Salary ranges are available for all positions at this location, taking into account roles with a comparable level of responsibility and impact in the relevant labor market and these salary ranges are regularly reviewed and adjusted in accordance with prevailing market conditions. The annualized base pay for the advertised position, located in the specified area, ranges from a minimum of $105,400 to a maximum of $198,100. The actual compensation is determined by various factors, including but not limited to the market pay for the jobs at each level, the responsibilities and skills required for each job, and the employee’s contribution (performance) in that role. To be considered within market range, a salary is at or above the minimum of the range. You may also have the opportunity to participate in discretionary equity (stock) based compensation and/or performance-based variable pay programs.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

承保运营负责人

Crum & ForsterUnited States$86,600 - $162,900/年Full Time今天
职能支持限定地区(需当地身份)

客户服务代表

Crum & ForsterUnited States$31,100 - $58,500/年Full Time今天
职能支持限定地区(需当地身份)

首席运营官,A&H

Crum & ForsterUnited StatesFull Time今天
职能支持限定地区(需当地身份)

← 返回全部职位