远程工作雷达

高级站点可靠性工程师

Senior Site Reliability Engineer

开发工程全球可投
公司Fingerprint
薪资未公开
工作地点不限地点
地域资格全球可投
时区要求无特别要求
用工类型未标注
发布时间昨天
数据来源Real Work From Anywhere
前往 Real Work From Anywhere 查看并投递 →
全球可投:该职位未限制候选人所在地区。仍需注意薪资可能按地区折算,以及实际签约方式(正式雇佣 / 独立合同)。

<div class="content-intro"><p><strong>Fingerprint</strong> 通过全球最精准的设备智能技术,帮助企业检测并阻止在线欺诈。我们凭借前沿的识别能力引领行业,并致力于将欺诈检测领域的创新想法和发现转化为现实。我们的客户包括创新型初创公司和领先的大型企业,如 Plaid、Dropbox 和 Booking.com。</p>
<p><strong>Fingerprint 是一家全球分布、100% 远程办公的公司。</strong> 我们入选了 2026 年福布斯最佳创业雇主榜单,并位列 2026 年 Inc. 5000 美国增长最快的私营企业第 803 名。</p>
<p><a href="https://www.crunchbase.com/organization/fingerprintjs">我们已融资 7700 万美元 </a>,并获得 Craft Ventures(<a href="https://www.tesla.com/">特斯拉</a>、<a href="https://facebook.com/">Facebook</a>、<a href="https://www.airbnb.com/">Airbnb</a>)、Nexus Venture Partners(<a href="https://www.postman.com/">Postman</a>、<a href="https://www.apollo.io/">Apollo.io</a>、<a href="https://min.io/">MinIO</a>、Druva)和 Uncorrelated Ventures(<a href="https://redis.io/">Redis</a>、<a href="https://rollbar.com/">Rollbar</a>、<a href="https://gradle.org/">Gradle</a>)的支持。</p>
<hr>
<p>&nbsp;</p>
<p>&nbsp;</p></div><p><strong>职位介绍</strong></p>
<p>你是否是一位系统思维的工程师,当生产环境告诉你一些你没想到的事情时,你会感到最开心?你更关心系统在凌晨 3 点面对未设计的负载时的表现,而不是它在图表上的样子?你是否希望负责一个每天处理数百万次身份识别请求的平台的可靠性,其中任何错误或延迟都会直接影响到客户?如果是这样,我们正好有适合你的机会。</p>
<p>我们正在寻找一名高级站点可靠性工程师加入我们的基础设施团队,负责平台在生产环境中的表现。这是一份需要亲自动手的工程岗位,而非监督性质——你将编写代码和基础设施,从头到尾负责系统,并根据你所负责的系统在我们不断增长的过程中是否保持快速、可用和可预测来衡量你的工作成果。</p>
<p>你将参与可观测性、事件响应、容量和性能、变更安全以及所有这些工作的工具链。你将定义你所负责关键路径的“可靠性”标准,对它们进行监控以便在客户之前发现问题,并与产品工程团队合作,使这一切变得常规化。

查看英文原文

<div class="content-intro"><p><strong>Fingerprint</strong> empowers enterprises to detect and stop online fraud with the world’s most accurate device intelligence.&nbsp; We lead our industry with bleeding-edge identification capabilities and work on turning new ideas and discoveries in the fraud detection space into reality. Our customers range from innovative startups to leading enterprise companies, including Plaid, Dropbox, and Booking.com.&nbsp;</p>
<p><strong>Fingerprint is a globally dispersed, 100% remote company.</strong> We were named on on the 2026 Forbes Best Startup Employers list and ranked #803 on the 2026 Inc. 5000 list of America’s fastest-growing private companies.&nbsp;</p>
<p><a href="https://www.crunchbase.com/organization/fingerprintjs">We have raised $77M </a>and are backed by Craft Ventures (<a href="https://www.tesla.com/">Tesla, </a><a href="https://facebook.com/">Facebook, </a><a href="https://www.airbnb.com/">Airbnb </a>), Nexus Venture Partners ( <a href="https://www.postman.com/">Postman</a>, <a href="https://www.apollo.io/">Apollo.io,</a> <a href="https://min.io/">MinIO</a>, Druva) and Uncorrelated Ventures ( <a href="https://redis.io/">Redis, </a><a href="https://rollbar.com/">Rollbar, </a><a href="https://gradle.org/">&nbsp;Gradle</a>).</p>
<hr>
<p>&nbsp;</p>
<p>&nbsp;</p></div><p><strong>About the role</strong></p>
<p>Are you a systems-minded engineer who is happiest when production tells you something you didn't expect? Do you care less about how a system looks on a diagram than about how it behaves at 3am under load it wasn't designed for? Do you want to own reliability for a platform that answers millions of identification requests a day, where being wrong or being slow is a customer-visible event? If so, we have the perfect opportunity for you.</p>
<p>We're looking for a Senior Site Reliability Engineer to join our Infrastructure team and take ownership of how our platform behaves in production. This is a hands-on engineering role, not an oversight one — you'll write code and infrastructure, own systems end to end, and be measured by whether the things you own stay fast, available, and predictable as we grow.</p>
<p>You'll work across observability, incident response, capacity and performance, change safety, and the tooling that makes all of it routine. You'll define what "reliable" means for the critical paths you own, instrument them so we know before customers do, and partner with product engineering teams to make their services operable by design rather than by heroics.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Own the reliability of core production systems end to end — you instrument them, set targets for them, operate them, and are accountable for how they behave under real traffic.</li>
<li>Define and maintain SLIs and SLOs for the critical paths you own, wire them into dashboards and alerts, and use error budget burn as the evidence base for what gets fixed next.</li>
<li>Drive alert quality: raise signal, kill noise, and close the gap where customers notice a problem before our monitoring does. Anomaly and correctness detection matter as much as uptime.</li>
<li>Take a lead role in incident response — investigate systematically across service boundaries, restore service, and write postmortems that produce follow-ups people actually complete.</li>
<li>Build secure, resilient, and cost-efficient infrastructure, with explicit attention to failure modes: timeouts and retries, backpressure and load shedding, graceful degradation, and blast radius containment.</li>
<li>Do capacity and performance work with real data — load testing, profiling, saturation analysis, and headroom planning ahead of growth rather than after an incident.</li>
<li>Improve change safety: progressive delivery, automated rollback, meaningful pre-production signal, and deployment practices that make shipping boring.</li>
<li>Manage infrastructure through code and configuration (we primarily use Terraform), consistently applying patterns that align with our overall service architecture.</li>
<li>Design, write, and ship software and developer-facing tooling that reduces toil and makes operating services straightforward for the engineers who own them.</li>
<li>Run deliberate failure testing — game days and chaos exercises, staging first — to find the gaps and safe limits before customers do.</li>
<li>Partner with product engineering teams on production readiness for new and high-risk services: capacity, failure modes, rollback plans, runbooks, and on-call handoff. Teach through review rather than gatekeeping.</li>
<li>Participate in the on-call rotation, and improve it: better runbooks, clearer escalation, less pager fatigue for everyone in it.</li>
<li>Approach all engineering work with a security lens — actively looking for vulnerabilities in your own work and in peer reviews.</li>
<li>Act as the go-to person for hard production problems in your area, and mentor engineers through code review, pairing, and design feedback so operational knowledge doesn't silo.</li>
</ul>
<p><strong>Qualifications</strong></p>
<ul>
<li>6–10 years of experience in SRE, production engineering, infrastructure, or backend engineering within primarily cloud-based environments (AWS preferred), with meaningful time spent responsible for systems in production.</li>
<li>A track record of owning a system end to end — you've designed something significant, shipped it, operated it, and lived with the consequences when it misbehaved.</li>
<li>Hands-on experience defining and operating against SLIs, SLOs, and error budgets — not just reading the book, but getting targets adopted and acted on.</li>
<li>Strong incident skills: you've led or been a primary responder on high-severity, customer-facing incidents, and you've improved how an organization learns from them.</li>
<li>Depth in distributed systems failure modes in high-throughput, low-latency environments — cache and database saturation, cascading failure, retry storms, capacity limits, degradation and load shedding.</li>
<li>Depth in cloud infrastructure fundamentals: networking, load balancing, containerization (EKS/Kubernetes), and distributed systems.</li>
<li>Strong hands-on experience managing infrastructure through code and configuration (Terraform or equivalent).</li>
<li>Solid programming skills in Go, Python, or a comparable language — you write real, production-ready software and can ship the fix rather than only recommend it.</li>
<li>Fluency with observability tooling (Datadog, Prometheus, Grafana, OpenTelemetry, or similar), including instrumenting systems yourself rather than inheriting dashboards.</li>
<li>Hands-on experience operating Redis/ElastiCache in production — including cluster/shard management, failover behavior, memory eviction policies, and scaling strategies. This is a current skill gap on our team, so depth here is a strong differentiator.</li>
<li>Fluency with software engineering best practices: source control, code review, comprehensive test coverage across edge cases and errors, and safe deployment.</li>
<li>A high level of personal ownership and autonomy, with real experience working without clearly defined requirements.</li>
<li>Pragmatism over purity — you know reliability competes with delivery, can make the case for the right investment at the right time, and can say when a risk is acceptable.</li>
<li>Strong written and verbal communication in English — clear technical design docs, PR reviews, incident updates, and postmortems that bring engineers outside your team along on a decision.</li>
<li>AI-native by default. You use AI tools as a normal part of how you investigate incidents, analyze telemetry, write runbooks, and build tooling — and you have opinions, from experience, about where they help and where they don't yet.</li>
</ul>
<p>&nbsp;</p>
<p><strong><em>Compensation &amp; Transparency</em></strong></p>
<p><em>For US-based employees, the cash compensation range for this role is $152,000 – $205,000. We set standard ranges for all US roles based on function, level, and geographic location, benchmarked against similar stage growth companies. To comply with local legislation and provide greater transparency, we share salary ranges on all job postings. However, these ranges are specific to the hiring location and may differ within or outside the US.&nbsp; Offers vary depending on, but not limited to, relevant experience, education, certifications/licenses, skills, training, and market conditions.</em></p>
<p>Due to regulatory and security reasons, there’s a small number of countries where we cannot have Fingerprint teammates based. Additionally, because Fingerprint is an all-remote company and people can join our workforce from almost any country, we do not sponsor visas. Fingerprint teammates need to be authorized to work from their home location.</p>
<p>We are dedicated to creating an inclusive work environment for everyone. We embrace and celebrate the unique experiences, perspectives and cultural backgrounds that each employee brings to our workplace. Fingerprint strives to foster an environment where our employees feel respected, valued and empowered, and our team members are at the forefront in helping us promote and sustain an inclusive workplace. We highly encourage people from underrepresented groups in tech to apply.</p>
<p>If you are applying as a resident of California, please read our CCPA notice<a href="https://dev.fingerprint.com/docs/dpa-ccpa"> here</a>.</p>
<p>If you are applying as a resident of the EU, please read our GDPR notice<a href="https://dev.fingerprint.com/docs/dpa-gdpr"> here.</a></p>
<ul>
<li>*<em>We have noticed a rise in recruiting impersonations across the industry, where scammers attempt to access candidates' personal and financial information through fake interviews and offers. All Fingerprint recruiting email communications will always come from the @fingerprint.com domain. Any outreach claiming to be from Fingerprint via other sources should be ignored.</em></li>
</ul><div class="content-conclusion"><p>Due to regulatory and security reasons, there’s a small number of countries where we cannot have Fingerprint teammates based. Additionally, because Fingerprint is an all-remote company and people can join our workforce from almost any country, we do not sponsor visas. Fingerprint teammates need to be authorized to work from their home location.</p>
<p>We are dedicated to creating an inclusive work environment for everyone. We embrace and celebrate the unique experiences, perspectives and cultural backgrounds that each employee brings to our workplace. Fingerprint strives to foster an environment where our employees feel respected, valued and empowered, and our team members are at the forefront in helping us promote and sustain an inclusive workplace. We highly encourage people from underrepresented groups in tech to apply.</p>
<p>If you are applying as a resident of California, please read our CCPA notice <a href="https://dev.fingerprint.com/docs/dpa-ccpa">here</a>.</p>
<p>If you are applying as a resident of the EU, please read our GDPR notice <a href="https://dev.fingerprint.com/docs/dpa-gdpr">here.</a></p>
<p><em>**We have noticed a rise in recruiting impersonations across the industry, where scammers attempt to access candidates' personal and financial information through fake interviews and offers. All Fingerprint recruiting email communications will always come from the @fingerprint.com domain. Any outreach claiming to be from Fingerprint via other sources should be ignored.</em></p></div>

本页面信息整理自 Real Work From Anywhere,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位