远程工作雷达

高级站点可靠性工程师 (SRE & 平台可靠性)

Senior Site Reliability Engineer (SRE & Platform Reliability)

开发工程限定地区(需当地身份)
公司Affirm
薪资未公开
工作地点Remote Spain
地域资格限定地区(需当地身份)
时区要求无特别要求
用工类型未标注
发布时间2026-08-07
数据来源Greenhouse
前往企业招聘页投递 →
注意地域限制:该职位明确限定在 Remote Spain 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

在Affirm,我们存在的意义是为那些重要的时刻提供清晰、可预测的支付方式,没有隐藏费用,没有意外,也不必在最重要的事情上做出妥协。

Affirm的站点可靠性工程(SRE)是一个规模虽小但至关重要的团队,帮助我们的工程合作伙伴以卓越的方式“运营他们所拥有的”,从而保护客户的体验。SRE通过定义框架和最佳实践来运营应用,构建工具,并提供培训和咨询服务来实现这一目标。SRE的许多职责包括:

  • 为团队和管理层提供应用性能的数据和可见性
  • 指导SLO的开发
  • 推动事件管理和分析流程
  • 引导变更管理和部署实践的实施
  • 参与服务和架构讨论
  • 建议可观测性和警报配置

SRE团队在多个领域拥有丰富的经验,包括:

  • 基础设施、平台和分布式系统
  • 容量管理、负载和混沌测试
  • 自动化、可观测性和配置管理
  • 开发和产品经验

SRE团队正在寻找有动力的软件和系统工程师,具备在Affirm工程组织内及之外构建、迭代和扩展事件生命周期、可靠性和弹性实践的经验。

你将负责:

  • 负责团队季度目标的制定和交付,带领团队中的工程师在不确定性中解决开放性问题,并确保在整个交付过程中每个人都得到支持。
  • 通过与基础设施、产品管理、开发者体验和分析团队协作,在产品开发周期中支持你的同事和利益相关者,参与构思、阐述技术限制,并在决策中合作,充分考虑风险和权衡。
  • 主动识别能够增强事件准备、响应和事后分析的技术解决方案和操作流程。
  • 通过创建和监控指标、在需要时进行升级,并支持“保持系统运行”和值班工作,来支持你团队的运维和可用性。
  • 通过为团队设定或改进代码审查和设计标准,并在团队之外倡导这些标准,来培养质量与责任的文化。
查看英文原文

At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most.

Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications, building tooling, and providing training and consulting. Some of the many SRE responsibilities are:

  • Providing data and visibility to teams and leadership on application performance
  • Guiding the development of SLOs
  • Driving the Incident Management and Analysis process
  • Steering the implementation of Change Management and Deployment practices
  • Engaging in service and architectural conversations
  • Recommending observability and alerting configurations

The SRE team benefits from experience across many domains including:

  • infrastructure, platform, and distributed systems
  • capacity management, load and chaos testing
  • automation, observability, and configuration management
  • development and product experience

The SRE team is seeking motivated software and systems engineers with the experience to build, iterate on, and expand incident lifecycle, reliability, and resilience practices throughout Affirms Engineering organization and beyond.

What You'll Do:

  • You will be responsible for owning and delivering quarterly goals for your team, leading engineers on your team through ambiguity to solve open-ended problems, and ensuring that everyone is supported throughout delivery.
  • You will support your peers and stakeholders in the product development lifecycle by collaborating with infrastructure, product management, developer experience & analytics by participating in ideation, articulating technical constraints, and partnering on decisions that properly consider risks and trade-offs.
  • You will proactively identify technical solutions and operational processes that strengthen incident readiness, response, and post-incident analysis.
  • You will support the operations and availability of your team’s artifacts by creating and monitoring metrics, escalating when needed, and supporting “keep the lights on” & on-call efforts.
  • You will foster a culture of quality and ownership on your team by setting or improving code review and design standards for your team, and advocating for them beyond your team through your writing and tech talks.
  • You will help develop talent on your team by providing feedback and guidance, and leading by example.
  • On-Call Rotation - There would be an on-call rotation for this role as a requirement.

What We Look For:

  • You have 4+ years of experience designing, developing and launching backend systems at scale using scripting and development languages like Bash, Python or Kotlin.
  • You have a track record of developing highly available distributed systems using technologies like AWS, MySQL and Kubernetes.
  • You have meaningful experience contributing in or driving parts of the Incident Lifecycle process, enabling actionable insights that improve the quality culture, reliability, resilience, and system performance.
  • You have 4+ years working in a Site Reliability or Production Engineering team
  • You demonstrate curiosity with empathy, and strong opinions loosely held
  • You have experience defining a technical plan for the delivery of a significant feature or system component with an elegant, simple and extensible design. You write high quality code that is easily understood and used by others.
  • You have experience in making impactful changes in a large code base, and have developed a suite of tools and practices that enable you and your team to do so safely.
  • Your experience demonstrates that you take ownership of your growth, proactively seeking feedback from your team, your manager, and your stakeholders.
  • You have strong verbal and written communication skills that support effective collaboration with our global engineering team.

Compensation & Benefits

Base Pay Grade - N

Equity Grade - 4

Employees new to Affirm typically come in at the start of the pay range. Affirm focuses on providing a simple and transparent pay structure which is based on a variety of factors, including location, experience and job-related skills.

Base pay is part of a total compensation package that may include monthly stipends for health, wellness and tech spending, and benefits (including 100% subsidized medical coverage, dental and vision for you and your dependents). In addition, the employees may be eligible for equity rewards offered by Affirm Holdings, Inc. (parent company).

ESP base pay range per year: €86,000  -  €122,000

Additional benefits include:

  • Flexible Spending Wallets for tech, food and lifestyle
  • Away Days - wellness days to take off work and recharge
  • Learning & Development programs
  • Parental benefit
  • Employee Resource & Community Groups

Location - Remote Spain

We are able to offer visa sponsorship for this role, but do require that someone is based in Spain for the role.

#LI-Remote

Remote-first with flexibility built in
Affirm is proud to be a remote-first company. Most roles can be done from almost anywhere within the country of employment. Some positions may occasionally require in-person work at an Affirm office, and a few are office-based due to the nature of the work. All new hires will be invited to attend an in-person onboarding experience.

Benefits designed for you
Our benefits reflect our commitment to care, transparency, and flexibility. Here are a few highlights:

  • Health coverage at no cost: We cover 100% of premiums for employees and their dependents.
  • Spending stipends: Monthly stipends support your tech setup, and the ability to choose health and wellness options that are right for you.
  • Time off to recharge: Flexible time off and generous holiday calendars help you rest when you need to.
  • Own a piece of what you build: Our employee stock purchase plan (ESPP) lets you buy Affirm stock at a discount.

We’re committed to providing an inclusive interview process, including accommodations for candidates with disabilities. If you need support, we’re happy to help.

For positions based in San Francisco or Los Angeles: Affirm considers qualified applicants with arrest and conviction records, as required by law.

By clicking "Submit Application," you acknowledge that you have read Affirm's Global Candidate Privacy Notice and consent to the use of your personal information as described.

本页面信息整理自 Greenhouse,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

软件工程师 I,前端

AffirmCanada101,000 - 151,000/年 CADpermanent4 天前
开发工程限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

← 返回全部职位