站点可靠性工程师
Site Reliability Engineer
职位概述
SingleStore 正在寻找一名站点可靠性工程师,帮助优化和扩展我们在三大主要云服务提供商上的托管服务。在这个职位中,你将处于前沿技术趋势的交汇点——一个高性能的分布式数据库,由 Kubernetes 管理,并在云中运行。这是一个很好的机会,通过以云为中心的 SRE 角色推动技术边界。
这是一项开发性工作,需要工程思维来解决运营挑战。你将加入一个全球分布的工程师团队,帮助推动公司内部的 SRE 实践。通过基础设施自动化,你将帮助我们在多个云平台上扩展我们的服务。这需要持续关注消除手动流程。你还将利用我们的监控平台,通过系统地识别和修复影响客户的任何问题,提升整体客户体验。作为 SRE,你还将协助诊断平台上的问题,借助对 SingleStore 查询引擎以及后端基础设施的深入理解。
你会做:
- 开发自动化平台,用于管理跨云服务商的基础设施部署
- 优化遥测平台,识别影响客户的事件,同时提供相关数据以支持调试
- 与工程团队合作,优化云架构下的服务性能
- 调试线上事件并进行后续的事故回顾和根本原因分析
- 参与以 SLA 为导向的值班轮班,包括非工作时间、周末和轮换节假日参与
你的经验:
- 基础设施自动化经验。熟悉 Python 和 Golang 更佳
- 熟悉 Kubernetes 和容器生态系统
- 强大的跨团队协作和沟通能力
- 熟悉至少一个云平台:AWS、Azure 或 Google Cloud
- 有调试、诊断和排查复杂生产软件的经验
- 计算机科学或相关领域的学士学位
关于我们:
SingleStore 以云原生数据库的速度和规模,为世界数据密集型应用提供支持。我们推出的分布式 SQL 数据库通过统一事务和分析,简化了您的数据架构,使 SingleStore 能够帮助数字领导者为客户带来卓越的实时数据体验。SingleStore 是一家获得风险投资的公司,总部位于旧金山,在 Sunnyvale 和 Rale 设有办公室
查看英文原文
Position Overview
SingleStore is seeking a Site Reliability Engineer to help optimize and scale our managed service offering across all three major cloud providers. In this role, you will be at the intersection of leading technology trends – A highly performant distributed database, managed by Kubernetes, running in the cloud. This is a great opportunity to push the boundaries with a cloud focused SRE role.
This is a development role, requiring an engineering mindset to solve operational challenges. You will be part of a globally distributed team of engineers, helping to drive SRE practices across the company. Through infrastructure automation, you will help us grow our service across multiple cloud platforms. This requires a relentless focus on eliminating manual processes. You will also leverage our monitoring platform to improve the overall customer experience by systematically identifying and fixing any issues impacting our customers. As an SRE, you will also help diagnose issues on the platform, leveraging a deep understanding of the SingleStore query engine along with the backend infrastructure.
What you'll do:
- Develop automation platform to manage infrastructure rollouts across cloud providers
- Optimize telemetry platform to identify customer impacting events while providing relevant data to drive debugging
- Partner with engineering team to optimize performance of services for cloud architecture
- Debug Live Site events and conduct follow-up postmortem and RCA analysis
- Participate in an SLA-driven on-call rotation, which will include after-hours, weekend, and rotating holiday participation.
Your Experience:
- Infrastructure automation experience. Python and Golang a plus.
- Knowledge of Kubernetes and the container ecosystem
- Strong cross group collaboration and communication skills
- Familiar with at least one of AWS, Azure, or Google Cloud
- Experience debugging, diagnosing and troubleshooting complex, production software
- B.S. Degree in Computer Science or related field
About us:
SingleStore delivers our cloud-native database with the speed and scale to power the world’s data-intensive applications. With a distributed SQL database that introduces simplicity to your data architecture by unifying transactions and analytics, SingleStore empowers digital leaders to deliver exceptional, real-time data experiences to their customers. SingleStore is venture-backed and headquartered in San Francisco with offices in Sunnyvale, Raleigh, Seattle, Boston, London, Lisbon, Bangalore, Dublin and Kyiv.
Consistent with our commitment to diversity & inclusion, we value individuals with the ability to work on diverse teams and with a diverse range of people.
To all recruitment agencies: SingleStore does not accept agency resumes. Please do not forward resumes to SingleStore employees. SingleStore is not responsible for any fees related to unsolicited resumes and will not pay fees to any third-party agency or company that does not have a signed agreement with the Company.
SingleStore values individuals for their unique skills and experiences, and we’re proud to offer roles in a variety of locations across the United States. Salary is based on permissible, non-discriminatory factors such as skills, experience, and geographic location, and is just one part of our total compensation and benefits package.
For candidates residing in California, please see our California Recruitment Privacy Notice. For candidates residing in the EEA, UK, and Switzerland, please see our EEA, UK, and Swiss Recruitment Privacy Notice.
Req ID: ENG00521
Originally posted on Himalayas