工程经理,SRE
Engineering Manager, SRE
关于远程办公
Remote 正在解决现代组织最大的挑战——轻松合规地处理全球雇佣。我们让各种规模的企业都能招聘、支付并管理国际团队。以我们的核心价值观为指导,以面向未来的工作文化为基础,我们的团队在全球范围内异步协作,致力于解决具有挑战性的问题。你可以发现 Remoters 在六大洲(南极洲即将加入!)工作,我们所有职位均为全远程。
以创新为核心价值观之一,我们已将自动化和人工智能能力融入每个职位的要求中。
我们鼓励 Remote 团队的每位成员带来自己的才能、经验和文化,帮助我们打造一流的 HR 平台。
如果你充满活力、好奇、有动力且有抱负,加入我们的世界吧。立即申请,定义工作的未来!
这份工作能为你提供什么
Remote 的 SRE 团队存在是为了让我们的工程师能够快速行动,客户获得稳定运行的产品。该团队负责 Kubernetes、AWS、PostgreSQL、CI 基础设施、我们的可观测性堆栈以及建立在这些之上的可靠性实践。
我们正在寻找一位团队负责人来带领这个团队。这是一个 60% 技术、40% 管理的角色。你将负责下属的职业发展,根据公司目标判断团队重点方向,并作为团队在工程部门的发言人。你还需要足够接近技术工作,以有说服力的方式设定方向,并在问题升级到你之前就意识到问题所在。
在 Remote,可靠性实践还在成长阶段而非成熟阶段。我们的 SLO 框架已在最初的几个团队上线,需要扩展到其余团队,如何平衡运营负载与项目交付还有大量实际工作要做。如果你希望加入一个基础已经建立,但有趣的问题仍然开放的团队,那么这就是你的团队。
你带来的
人员领导力
- 你曾领导过 SRE、基础设施或平台工程团队,并且关注下属的成长、绩效和职业发展,而不仅仅是他们的迭代。
- 你既辅导技术技能也辅导软技能,你能指出因你而成长的人。
- 你能够直接且早期地处理表现不佳的情况,做到清晰且富有同理心。
- 你曾招聘过工程师,并能区分一次好的面试和一名优秀的工程师。
- 你善于观察团队动态,并能有效解决冲突。
查看英文原文
About Remote
Remote is solving modern organizations’ biggest challenge – navigating global employment compliantly with ease. We make it possible for businesses of all sizes to recruit, pay, and manage international teams. With our core values at heart and future focused work culture, our team works tirelessly on ambitious problems, asynchronously, around the world. You can find Remoters working from 6 different continents (Antarctica left to go!) and all of our positions are fully remote.
With Innovation as one of the core values, we have built Automation and AI capabilities into the requirements for every role.
We encourage every member of the Remote team to bring their talents, experiences and culture to the table to help us build the best-in-class HR platform.
If you are energetic, curious, motivated and ambitious, be part of our world. Apply now and define the future of work!
What this job can offer you
Remote's SRE team exists so that our engineers can move quickly and our customers get a product that stays up. The team owns Kubernetes, AWS, PostgreSQL, CI infrastructure, our observability stack and the reliability practices that sits on top of all of it.
We are looking for a Team Leader to run that team. This is a 60% IC, 40% leadership role. You will own the career development of your reports, steer the teams focus using judgment against the company goals, and you will be the spokesperson for the team across engineering. You will also stay close enough to the technical work to set direction with credibility and to know when something is going wrong before it is escalated to you.
Reliability practice at Remote is maturing rather than mature. Our SLO framework is live on its first few teams and needs to reach the rest, there is real work to do on how we balance operational load against project delivery. If you want a team where the foundations are in place and the interesting problems are still open, this is that team.
What you bring
People leadership
- You have led an SRE, infrastructure or platform engineering team, and you have owned your reports' growth, performance and career progression rather than just their sprints.
- You coach both craft and the soft skills, you can point to people who grew because of it.
- You handle underperformance directly and early, with clarity and empathy.
- You have hired engineers, and you can tell the difference between a good interview and a good engineer.
- You read team dynamics well and you resolve conflict rather than routing around it.
- You have a natural talent fostering commitment to the goals of the company.
Technical depth
- A hands-on background in site reliability, DevOps or cloud infrastructure engineering, deep enough that you can review your team's work, challenge a design and be taken seriously in an incident.
- Kubernetes in production, including the operational reality of it rather than the happy path.
- AWS at meaningful scale.
- Hands on AI building, enablement, scaling AI infrastructure.
- Solid o11y practices and principles,
- Infrastructure as code with Terraform.
- CI/CD systems such as GitLab CI, GitHub Actions or Jenkins.
- Docker and shell scripting.
- You have run a reliability practice: incident response, on-call, SLOs and error budgets, and the discipline of turning incidents into changes that stick.
- Understanding and history of working in regulated environments.
Ways of working
- You prioritise exceptionally well when operational load and project work compete, and you protect your team's focus without dropping the operational commitment.
- You write clearly. Remote is fully distributed and async, so most of your leadership will happen in writing.
- You build relationships across teams. A lot of SRE's value comes from being the team others bring problems to early.
Nice to have
- Working knowledge of a backend language, ideally Elixir, or otherwise Java, Clojure, Node.js, Python or similar.
- Depth in modern observability: OpenTelemetry, distributed tracing, and tools such as Honeycomb.
- Database operations experience, particularly PostgreSQL or Aurora performance, connection pool health and query tuning.
- Running and configuring Linux systems outside a cloud environment.
- Security capability from both a defensive and an offensive standpoint.
- Cloud cost management and FinOps.
- Experience growing a team from a small base, including building the hiring bar as you go.
Key Responsibilities
Your people
- The full career lifecycle of your reports: onboarding, feedback, performance assessment, progression and hiring.
- Team health, dynamics and the retrospective habit that keeps them honest.
- Being the team's spokesperson to the rest of engineering and to senior leadership.
Delivery
- The SRE goals: what the team commits to, in what order, and why.
- The support rotation and on-call model.
The platform
- Remote's core infrastructure: Kubernetes, AWS, PostgreSQL, DNS and TLS, CI infrastructure.
- The reliability practice: SLOs, error budgets, incident response and the observability stack.
- The partnership with our Security team on threats, patching and infrastructure controls, including our audit and compliance obligations.
- The vendor relationships that sit behind the platform, including renewals and commercial conversations with support from your Director.
Practicals
- You will report to: Director of Engineering, Platform
- Team: Site Reliability Engineering, part of Platform Engineering
- Direct reports: 4
- Location: Anywhere in the world. Our current coverage is strongest in EMEA and APAC, so candidates who overlap with either, or who can help us close the Americas gap, are especially welcome.
- Start date: As soon as possible
Application process
- Interview with recruiter
- Interview with hiring manager
- Scenario interview with a peer team leader
- Executive interview
- Bar Raiser interview
- Offer + Prior employment verification check
Remote's Total Rewards philosophy is to ensure fair, unbiased compensation and fair equity pay along with competitive benefits in all locations in which we operate. We do not agree to or encourage cheap-labor practices and therefore we ensure to pay above in-location rates. We hope to inspire other companies to support global talent-hiring and bring local wealth to developing countries.
At first glance our salary bands seem quite wide - here is some context. At Remote we have international operations and a globally distributed workforce. We use geo ranges to consider geographic pay differentials as part of our global compensation strategy to remain competitive in various markets while we hiring globally.
Our salary ranges are determined by role, level and location, and our job titles may span more than one career level. The actual base pay for the successful candidate in this role is dependent upon many factors such as location, transferable or job-related skills, work experience, relevant training, business needs, and market demands. The base salary range may be subject to change.
At Remote, we foster internal mobility as a key element of our culture of employee growth and development, supported by a compensation philosophy that guarantees pay equity and fairness. Therefore, all compensation changes associated with an internal move will be reviewed by the Total Rewards & People Enablement team on a case by case basis.
The annual salary range for this full-time position is
$75,450—$169,700 USD
Benefits
Our full benefits & perks are explained in our handbook at . As a global company, each country works differently, but some benefits/perks are for all Remoters:
- work from anywhere
- flexible paid time off
- flexible working hours (we are async)
- 16 weeks paid parental leave
- mental health support services
- stock options
- learning budget
- home office budget & IT equipment
- budget for local in-person social events or co-working spaces
How you’ll plan your day (and life)
We work async at Remote which means you can plan your schedule around your life (and not around meetings). Read more at .
You will be empowered to take ownership and be proactive. When in doubt you will default to action instead of waiting. Your life-work balance is important and you will be encouraged to put yourself and your family first, and fit work around your needs.
If that sounds like something you want, apply now!
How to apply
- Please fill out the form below and upload your CV with a PDF format.
- We kindly ask you to submit your application and CV in English, as this is the standardised language we use here at Remote.
- If you don’t have an up to date CV but you are still interested in talking to us, please feel free to add a copy of your LinkedIn profile instead.
We will ask you to voluntarily tell us your pronouns at interview stage, and you will have the option to answer our anonymous demographic questionnaire when you apply below. As an equal employment opportunity employer it’s important to us that our workforce reflects people of all backgrounds, identities, and experiences and this data will help us to stay accountable. We thank you for providing this data, if you chose to.
Please note we accept applications on an ongoing basis.
Originally posted on Himalayas