站点可靠性工程师
Site Reliability Engineer
## 关于 Prove
随着世界向以移动设备为主的经济转型,企业需要现代化其获取、与消费者互动和赋能的方式。Prove 的以电话为中心的身份令牌化和被动加密认证解决方案减少了摩擦,增强了所有数字渠道的安全性、隐私性,并加速了收入,同时降低了运营成本和欺诈损失。超过 1000 家企业客户使用 Prove 的平台,每年在包括银行、贷款、医疗保健、游戏、加密货币、电子商务、市场平台和支付在内的各个行业处理 200 亿次客户请求。有关 Prove 的最新动态,请[在 LinkedIn 上关注我们](https://www.linkedin.com/company/proveidentity/)。
Prove 正在推动数字身份的未来。我们正在寻找能够产生影响的 Provers。我们谈论的是那些能够自主行动、在快节奏环境中茁壮成长、快速处理信息并做出明智决策的专业人士。这项工作充满挑战,不仅需要聪明,还需要自然的好奇心和坚持。团队合作对我们也很重要——我们共同工作,也一起娱乐。
Prove 有宏伟的计划,我们对未来发展充满期待。如果这听起来像适合你的地方——加入我们的团队吧!
**职位:高级站点可靠性工程师**
**部门:平台工程**
**汇报对象:平台工程总监**
**FLSA 状态:非豁免**
**地点:美国远程办公**
**职位简介**
我们正在寻找中高级站点可靠性工程师加入我们的平台工程团队。在这个职位上,你将负责设计、实施、维护和部署高可用性、复杂、可扩展且可靠的系统,利用自动化、有效的监控和基础设施即代码。与我们的应用工程团队紧密合作,确保我们的服务达到最高标准的可靠性、性能和安全性。
**高级职位的主要职责**
Prove 的站点可靠性工程团队负责确保现有和开发中的产品的最大正常运行时间。合格的候选人应熟悉方法之间的差异以及对结果的责任,并能够展示和记录其相关经验。
**可观测性领导力**
- 设计并实现跨我们的基础设施和应用程序的全面可观测性解决方案
- 建立指标、日志和追踪系统,以快速识别
查看英文原文
## About Prove
As the world moves to a mobile-first economy, businesses need to modernize how they acquire, engage with and enable consumers. Prove’s phone-centric identity tokenization and passive cryptographic authentication solutions reduce friction, enhance security and privacy across all digital channels, and accelerate revenues while reducing operating expenses and fraud losses. Over 1,000 enterprise customers use Prove’s platform to process 20 billion customer requests annually across industries, including banking, lending, healthcare, gaming, crypto, e-commerce, marketplaces, and payments. For the latest updates from Prove, [follow us on LinkedIn.](https://www.linkedin.com/company/proveidentity/)
Prove is driving the future of digital identity. We are looking for Provers who know how to make an impact. We’re talking self-starting professionals who thrive in a fast-paced environment, process information quickly, and make intelligent decisions. The work is challenging and requires not only smart but natural curiosity and tenacity. Teamwork is also important to us – we work together and play together.
Prove has big plans, and we’re excited about the future. If this sounds like the place for you – come join our team!
**Title: Site Reliability Engineer, Senior Site Reliability Engineer**
**Department: Platform Engineering**
**Reports To: Director, Platform Engineering**
**FLSA Status: Exempt**
**Location: US Remote**
**Job Summary**
We are seeking Mid to Senior level Site Reliability Engineers to join our Platform Engineering team. In this role, you will be instrumental in designing, implementing, maintaining and deploying highly available complex, scalable and reliable systems leveraging automation, effective monitoring and infrastructure-as code. Working closely with our application engineering teams to ensure our services meet the highest standards of reliability, performance, and security.
**Key Responsibilities for Senior level**
The Site Reliability Engineering teams at Prove are responsible for driving maximum uptime for existing and developing products. Qualified candidates will be well versed in the difference between methods and ownership of outcomes and be able to demonstrate and document their relevant experience.
**Observability Leadership**
- Design and implement comprehensive observability solutions across our infrastructure and within applications
- Establish metrics, logging, and tracing systems that enable quick identification and resolution of issues
- Create alerting thresholds and automated responses based on service level objectives (SLOs)
- Provide actionable insights into service to service communications
**Infrastructure Management**
- Design, build, and maintain scalable cloud infrastructure on AWS
- Implement infrastructure-as-code using tools such as Terraform
- Automate routine operational tasks to reduce toil and improve efficiency
- Ensure infrastructure security compliance and implement least-privilege access controls
- Design and implement infrastructure-as-code deployments for container based applications
- Scale containers based on custom metrics for applications and critical observability infrastructure
**Incident Response**
- Conduct thorough post-incident reviews and implement preventative measures
- Use observability data to perform root cause analysis and system improvements
- Participate in a 24/7 on call rotation to achieve 99.999% system availability.
**Required Qualifications for Senior level**
- 5+ years of experience in Site Reliability Engineering, Platform Engineering or equivalent experience. Software Engineering roles with a strong infrastructure and production engineering aspect also qualify.
- Expert knowledge of observability platforms and practices (OpenTelemetry, Prometheus, Grafana, Jaeger, ELK stack / Splunk, etc)
- Experience with Kubernetes and container orchestration
- Strong experience with infrastructure-as-code tools (Terraform, Spacelift, Pulumi)
- Proficiency in at least one programming language ( Go, Python )
- Deep understanding of cloud platforms, preferably AWS
- Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
**Key Responsibilities for IC3**
The Site Reliability Engineering teams at Prove are responsible for driving maximum uptime for existing and developing products. Qualified candidates will be well versed in the difference between methods and ownership of outcomes and be able to demonstrate and document their relevant experience.
**Preferred Qualifications**
- Experience with distributed systems and microservice architectures
- Experience working in a high compliance environment
- Hand-on experience instrumenting code with OpenTelemetry
- Familiarity with service mesh technologies
- Contributions to open-source projects
- Experience in the identity verification or financial technology industry
- Application development experience
**Optimize**
- Improve new and existing systems by increasing reliability, performance, and scalability
- Automate routine operational tasks to reduce toil and improve efficiency
- Ensure infrastructure security compliance and implement least-privilege access controls
- Implement efficient infrastructure that balances rapid development and cost
- Embrace technological changes and development practices while maintaining reliability
**Respond**
- Participate in a 24/7 on-call rotation
- Conduct thorough post-incident reviews and implement preventative measures
- Use observability data to identify system improvements
**Run**
- Implement infrastructure as code in a myriad of high compliance development, production, and other environments
- Scale developer experiences by being the standard bearer of an opinionated platform approach
**Required Qualifications for IC3**
- 3+ years of experience in Site Reliability or Platform Engineering teams
- Deep understanding of cloud platforms, particularly AWS
- Strong experience with Kubernetes and container orchestration
- Experience withTerraform and infrastructure-as-code tools
- Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
**Preferred Qualifications**
- Experience with distributed systems and microservice architectures
- Experience working in a high compliance environment
- Experience with holistic monitoring and alerting for developing platforms
- Skilled proficiency in at least one programming language (Go, Python)
**Benefits & Perks for FTE Provers:**
- Competitive salaries & Bonus Plan (for eligible roles) and Equity Plan
- Modern Health for financial, mental, and physical wellness
- 401(k) Retirement Plan & Match (US Offices) and Local Country Pension (International Offices)
- Unlimited Vacation and Flexible hours
- Comprehensive medical benefits for you and your family ❤️
- Emotional & Physical Wellness – Access to wellness services (EAP & Prove Well-Being Reimbursement)
- Bottomless snacks & beverages for certain office locations
- Daily GrubHub stipend for lunch if coming into the office (US Offices)
- A great place to work and connect with other talented Provers like yourself!
This position description should not be considered the final description of the position. The position description is not intended to be an all-inclusive list of duties and standards of the positions. It should be assumed that we would, to some extent, structure responsibilities in accordance with the successful candidate’s capabilities and changing business conditions. Incumbents will follow any other instructions, and perform any other related duties, as assigned by their supervisor.
**Site Reliability Engineer:**
**Metro 2: $130,000 - 150,000**
**Metro 3: $120,000 - 135,000**
**Senior Site Reliability Engineer:**
**Metro 2: $166,000 - 185,000**
**Metro 3: $153,000 - 171,000**
Plus variable commission / company bonus. Offered salary will be determined by the applicant’s education, experience, knowledge, skills, geo-location and abilities, as well as internal equity and alignment with market data.
Prove follows a market driven compensation philosophy based on geographic location and respective market rates. Job offers will be aligned to location. Please speak with your recruiter if you have questions. Prove defines:
- Metro 2 - NYC metro area, Seattle metro area, Los Angeles metro area, and the Miami metro area.
- Metro 3 - all other cities across the domestic United States, with the exception of the San Francisco Bay Area.
**Benefits & Perks for FTE Provers:**
- Competitive salaries & Bonus Plan (for eligible roles) and Equity Plan
- Modern Health for financial, mental, and physical wellness
- 401(k) Retirement Plan & Match (US Offices) and Local Country Pension (International Offices)
- Unlimited Vacation and Flexible hours
- Comprehensive medical benefits for you and your family ❤️
- Emotional & Physical Wellness – Access to wellness services (EAP & Prove Well-Being Reimbursement)
- Bottomless snacks & beverages for certain office locations
- Daily GrubHub stipend for lunch if coming into the office (US Offices)
- A great place to work and connect with other talented Provers like yourself!
Don’t meet every single requirement? Studies have shown that women and people of color are less likely to apply to jobs unless they meet every single qualification. At Prove we are dedicated to building a diverse, inclusive and authentic workplace, so if you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyways. You may be just the right candidate for this or other roles.
**Equal Opportunity Employment:**
Prove is an equal opportunity employer committed to providing equal employment opportunity for all people regardless of race, color, religion, gender or sexual orientation, age, marital status, national origin, citizenship status, disability, veteran status or other personal characteristics
**Privacy & Data Protection:**
When you are applying for a job at Prove, we collect and use your personal information in the job application process. To understand more about how Prove uses your personal information, please see our [Recruitment Privacy Policy](https://www.prove.com/legal/recruitment-privacy-notice) on our website.