远程工作雷达

高级软件工程师 - SRE

Senior Software Engineer - SRE

开发工程限定地区(需当地身份)
公司Mercury
薪资未公开
工作地点San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States
地域资格限定地区(需当地身份)
时区要求无特别要求
用工类型未标注
发布时间7 天前
数据来源Greenhouse
前往企业招聘页投递 →
注意地域限制:该职位明确限定在 San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

当英格兰埃克斯穆尔国家公园内的塔尔步道(Tarr Steps)这座由巨石搭建的步行桥在1942年的一场洪水中被冲毁后,皇家工程兵重新修建了它。但1952年它再次被冲毁。于是皇家工程兵又添加了更多石头:如此反复,就像西西弗斯式的荒诞教训。

当然,他们这么做是因为塔尔步道是一个历史遗迹,据信建于青铜时代。现代桥梁建造方式大不相同,通常维护需求也少得多。虽然我们欣赏古老事物的魅力,但Mercury正在打造银行的未来。"古雅"和"过时"并不是我们系统所追求的价值。重启一台崩溃的服务器并希望请求洪流不会将其冲走,这只能作为应急策略,而我们更积极地寻求持久的解决方案,而不是单调的运维工作。你将带来这种理念,并教会其他人如何践行。

到目前为止,Mercury的稳定性团队主要构建了产品团队可以采用的平台级组件。然而这些团队希望我们更紧密地与他们合作,以成熟他们的实现和实践。我们正在创建一个SRE团队,将轮换在各个产品团队中工作。你将理解该团队的领域,识别可靠性、可观测性、性能和准备度方面的改进机会,并创造自我强化的良性循环。

作为该职位的一部分,你将:

  • 与产品团队紧密合作,通过建立和优化值班、监控、警报和操作手册等实践,帮助他们提升运维成熟度
  • 定期与产品团队进行演练日活动,帮助他们更充分地准备调查和快速处理事件
  • 深入Haskell和TypeScript编写的应用程序代码,实现可靠性技术,如重试、更好的错误处理、更好的日志记录、断路器等
  • 引导服务级别目标(SLOs)朝着产品团队负责的有意义的客户结果发展。这些SLOs将成为产品是否按预期工作的强大信号
  • 在设计文档和代码审查中倡导可靠性实践
  • 识别阻碍调试、事件响应和业务智能的可观测性缺口,并协助解决这些问题
  • 推动非产品工程团队可以推动的长期改进
  • 在嵌入产品团队或作为更广泛的工程轮班的一部分时,参与产品团队的值班轮班,帮助我们改进流程和学习方式
查看英文原文

When the Tarr Steps, a footbridge assembled of heavy stones in Exmoor National Park in England, washed away in a flood in 1942, the Royal Engineers rebuilt it. Then it washed away again in 1952. So the Royal Engineers heaved in more stones: on and on, like a Sisyphean lesson in absurdity.

Of course, they do this because the Tarr Steps are a historical monument, believed to be built during the Bronze Age. Modern bridge building looks a lot different and generally requires less upkeep. While we appreciate the charm of ancient things, Mercury is engineering the future of banking*. "Quaint" and "archaic" are not values we seek in our systems. Rebooting a tumbling server and hoping a flood of requests doesn't wash it away can be an emergency tactic, but we actively seek out durable solutions over monotonous ops work. You'll bring this ethos and teach others how to live it too.

Up to this point, the Stability team at Mercury has primarily built platform-level constructs that product teams adopt. However, those teams are asking for us to work more closely with them to mature their implementations and practices. We are creating an SRE team that will rotate through product teams. You will understand the team's domain, identify opportunities for improvements in reliability/observability/performance/preparedness and create self-reinforcing, virtuous cycles.

As part of this role, you will:

  • Embed with product teams, helping them improve their operational maturity by setting up and refining practices around on-call, monitoring, alerting, and run books
  • Run regular game day exercises with product teams, helping them feel more prepared to investigate and quickly remediate incidents
  • Jump into application code written in Haskell & TypeScript and implement reliability techniques such as retries, better error handling, better logging, circuit breaking, etc
  • Steer SLOs towards meaningful customer outcomes that product teams are accountable for. Those SLOs become a strong signal for whether the product is working as intended
  • Champion reliability practices through design document and code reviews
  • Identify observability gaps that hinder debugging, incident response, and business intelligence and help close those
  • Advocate for longer-term improvements that non-product engineering teams can drive
  • Participate in the product team's on-call rotation while embedding or as part of a more general engineering rotation, helping us improve processes and how we learn from incidents

The ideal candidate for the role:

  • Has past Site Reliability Engineering or DevOps experience
  • Has measurable examples of influencing an organization towards greater reliability
  • Has significant experience with PostgreSQL
  • Has authored and operated Temporal workflows
  • Has experience with observability platforms like Grafana or Honeycomb
  • Has familiarity with OpenTelemetry

If this role interests you, we invite you to explore our public demo at personal-demo.mercury.com.

*Mercury is a fintech company, not an FDIC-insured bank. Banking services provided through Choice Financial Group and Column N.A., Members FDIC.

Mercury values diversity & belonging and is proud to be an Equal Employment Opportunity employer. All individuals seeking employment at Mercury are considered without regard to race, color, religion, national origin, age, sex, marital status, ancestry, physical or mental disability, veteran status, gender identity, sexual orientation, or any other legally protected characteristic. We are committed to providing reasonable accommodations throughout the recruitment process for applicants with disabilities or special needs. If you need assistance, or an accommodation, please let your recruiter know once you are contacted about a role.

#LI-GC1
Total Rewards
The total rewards package at Mercury includes base salary, equity (stock options/RSUs), and benefits.

Our salary and equity ranges are highly competitive within the SaaS and fintech industry and are updated regularly using the most reliable compensation survey data for our industry. New hire offers are made based on a candidate’s experience, expertise, geographic location, and internal pay equity relative to peers.

Our target new hire base salary ranges for this role are the following:

US employees (any location):
$200,700—$250,900 USD

Canadian employees (any location):
$189,700—$237,100 CAD

本页面信息整理自 Greenhouse,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

业务风险官

MercurySan Francisco, CA, New York, NY, Portland,今天
职能支持限定地区(需当地身份)

交易运营负责人

MercurySan Francisco, CA, New York, NY, Portland,昨天
职能支持限定地区(需当地身份)

反洗钱调查员

MercuryUnited States$109,500 - $152,100/年permanent4 天前
职能支持限定地区(需当地身份)与中国几乎无重叠,需长期倒时差

← 返回全部职位