高级站点可靠性工程师,tvScientific
Sr. Site Reliability Engineer, tvScientific
关于Pinterest:
全球数百万人来到我们的平台,寻找创意灵感,畅想新的可能性,并规划将伴随一生的回忆。在Pinterest,我们致力于为每个人带来创造理想生活的灵感,而这一切始于产品背后的每一个人。
在这里开启职业生涯,你将点燃创新,将热情转化为增长机会,庆祝彼此的独特经历,并拥抱灵活的工作方式,发挥最佳水平。打造你热爱的职业?一切皆有可能。
在Pinterest,AI不仅仅是一个功能,它是一个强大的合作伙伴,增强我们的创造力并放大我们的影响力,我们正在寻找渴望成为其中一员的候选人。为了全面了解你的经验和能力,我们将考察你的基础技能以及你与AI协作的方式。
通过我们的面试流程,最重要的是你能清晰地解释你的方法,不仅展示你掌握的知识,更展现你的思维方式。你可以在此处了解更多关于我们的AI面试理念以及我们如何在招聘过程中使用AI的信息。
关于tvScientific
tvScientific是首个且唯一的CTV广告平台,专为效果营销人员打造。我们利用海量数据和前沿科学,自动化并优化电视广告以推动业务成果。我们的解决方案将媒体购买、优化、测量和归因整合在一个高效平台上。我们的平台由行业领袖打造,他们拥有程序化广告、数字媒体和广告验证领域的长期经验,现在专门构建了一个广告商可以信赖的CTV效果平台,助力其业务增长。
我们正在寻找一位高级站点可靠性工程师,帮助运营、扩展并持续改进基于AWS、Kubernetes/EKS和ArgoCD驱动的GitOps工作流的云原生平台。该职位将在提升基础设施和交付生态系统的可靠性、可扩展性、自动化、可观测性和运营成熟度方面发挥关键作用。
理想的候选人是一位高度动手能力强的工程师,具备丰富的生产环境经验,并有成功使用基础设施即代码、自动化和现代Kubernetes操作实践构建和维护可靠平台的能力。
你将负责:
- 确保生产基础设施和平台服务的可靠性、可用性和性能
- 运营和扩展Kubernetes平台,包括go
查看英文原文
About Pinterest:
Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product.
Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible.
At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI.
Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here.
About tvScientific
tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.
We are seeking a Senior Site ReliabilityEngineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows. This role will be instrumental in advancing the reliability, scalability, automation, observability, and operational maturity of our infrastructure and delivery ecosystem.
The ideal candidate is a highly hands-on engineer with strong production experience and a proven ability to build and support resilient platforms using infrastructure as code, automation, and modern Kubernetes operational practices.
What you’ll do:
- Ensuring the reliability, availability, and performance of production infrastructure and platform services
- Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads
- Managing GitOps-based deployment workflows using ArgoCD and Helm
- Driving infrastructure provisioning and change management through Terraform/Terragrunt
- Building and supporting CI/CD automation and deployment workflows using GitHub Actions
- Leading incident response efforts, root cause analysis, and post-incident improvement initiatives
- Reducing operational toil through scripting, tooling, and process automation
- Advancing observability practices across logs, metrics, traces, dashboards, and alerting
- Supporting secure secrets integration, IAM-aware operations, and platform guardrails
- Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes
What we're looking for:
- 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
- Strong hands-on experience operating AWS in production environments
- Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration
- Proven experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns
- Experience implementing and operating ArgoCD within a GitOps delivery model
- Strong hands-on experience with Helm
- Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management
- Solid scripting and automation skills using Bash and/or Python
- Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions
- Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems
- Experience with monitoring, alerting, and observability in production environments
- Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow-through after outages
- Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams
- Bachelor’s degree in computer science, engineering, a related field or equivalent experience
- Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs
- Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review)
- High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables.
In-Office Requirement Statement:
- We recognize that the ideal environment for work is situational and may differ across departments. What this looks like day-to-day can vary based on the needs of each organization or role.
Relocation Statement:
- This position is not eligible for relocation assistance. Visit our PinFlex page to learn more about our working model.
#LI-SM4
#LI-REMOTE
At Pinterest we believe the workplace should be equitable, inclusive, and inspiring for every employee. In an effort to provide greater transparency, we are sharing the base salary range for this position. The position is also eligible for equity. Final salary is based on a number of factors including location, travel, relevant prior experience, or particular skills and expertise.
Information regarding the culture at Pinterest and benefits available for this position can be found here.
US based applicants only
$139,764—$287,749 USD
Our Commitment to Inclusion:
Pinterest is an equal opportunity employer and makes employment decisions on the basis of merit. We want to have the best qualified people in every job. All qualified applicants will receive consideration for employment without regard to race, color, ancestry, national origin, religion or religious creed, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, age, marital status, status as a protected veteran, physical or mental disability, medical condition, genetic information or characteristics (or those of a family member) or any other consideration made unlawful by applicable federal, state or local laws. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. If you require a medical or religious accommodation during the job application process, please complete this form for support.
By submitting this application, I certify that all information submitted in my application and throughout the hiring process is true, accurate, and complete to the best of my knowledge. I understand that any false statement, omission, or misrepresentation may disqualify me from employment consideration or result in termination if discovered after hire.