高级站点可靠性工程师
Senior Site Reliability Engineer
Intuition Machines 使用 AI/ML 构建企业安全产品。我们将研究成果应用于服务数亿人的系统,团队分布在全球各地。你可能已经了解我们最知名的产品——hCaptcha 安全套件。我们的方法很简单:低开销、小团队和快速迭代。
作为高级站点可靠性工程师,你将专注于与性能、可用性、安全性和成本效益相关的工程解决方案。我们认为这些非功能性特性是我们和客户的核心需求。你将在我们互联网规模系统的多个层面(基础设施、数据、应用逻辑)工作,并构建解决方案。
使用 AI:代码代理无疑是很有用的工具。我们提供对前三大模型的访问,并且是最早采用评估优先开发流程的公司之一。熟悉使用代理进行编码是所有面试的一部分。然而,可靠性和正确性对我们来说至关重要。你需要阅读并理解每行你署名的代码,并且它将由人和机器共同审查。
你将负责:
- 与大规模系统合作(每秒处理数百万个请求,为数百万用户提供服务,跨多个云服务商)。
- 开发提升性能、可用性、安全性和成本效益的解决方案。
- 确保系统稳定、快速,并保持开发团队的生产力,确保每次同行发布都能在质量、安全性、正常运行时间、交付速度、威胁检测和客户参与度等方面提升整体性能。
- 从客户、内部社区、新旧系统指标中获取改进想法、优先级和能力。快速做出决策。
- 具有创造力,渴望在一个可以直接创造价值的环境中工作,并成为改善客户体验的推动力。
我们寻找:
- 精通 Kubernetes。
- 精通应用程序、基础设施和网络的监控。
- 软件工程背景,精通基于 Kubernetes 的系统中的后端开发。
- 在以下一种或多种语言中具有强大的编程技能:Python、JavaScript、Go、C++、Rust。
- 对网络、代理、内容分发网络(Cloudflare)有深入的理解和实践经验。
- 多云经验,包括虚拟网络、负载均衡、Web 应用防火墙。
- 强大的 CI/CD 经验。
- 具有开发和运维的实际经验。
查看英文原文
Intuition Machines uses AI/ML to build enterprise security products. We apply our research to systems that serve hundreds of millions of people, with a team distributed around the world. You are probably familiar with our best-known product, the hCaptcha security suite. Our approach is simple: low overhead, small teams, and rapid iteration.
As a Senior Site Reliability Engineer, you will focus on engineering solutions related to performance, availability, security, and cost-effectiveness. We consider these non-functional features to be core requirements for us and our customers. You will work at multiple layers of our internet-scale system (infrastructure, data, application logic) and build the solutions.
Using AI: Coding agents are indisputably useful tools. We provide access to the top 3 models, and were early adopters of evals-first development flows. Familiarity with coding using agents is part of all interviews. However, reliability and correctness are critical for us. You will need to read and understand every line of code with your name on it, and it will be reviewed by both people and machines.
What you will do:
- Work with large-scale systems (handling millions of requests per second, serving millions of users, across multiple cloud providers).
- Develop solutions to enhance performance, availability, security, and cost-effectiveness.
- Keep us up, keep us fast, and keep our dev teams productive ensuring that every peer release improves performance across the spectrum including quality, security, uptime, speed-to-deliver, threat detection, and customer engagement.
- Source improvement ideas, priority and capabilities from customers, the internal community, new and existing system metrics. Make decisions rapidly.
- Be creative and desire an environment where you can directly create value and be a force to improve the experience for our customers.
What we are looking for:
- Expert in Kubernetes.
- Expert in monitoring applications, infrastructure and network.
- Background in software engineering with expertise in backend development within Kubernetes-based systems.
- Strong programming skills in one or more of the following languages: Python, JavaScript, Go, C++, Rust.
- Strong understanding and experience in networking, proxies, content delivery networks (Cloudflare)
- Multi cloud experience including virtual networking, load balancing, web application firewall.
- Strong experience with CI/CD.
- Hands-on experience in development and orchestration within high-scale, high-uptime, and high-reliability environments.
- Minimum of six years of hands-on experience in related roles (engineering, DevOps, SRE).
- Familiarity with distributed systems, including queue-first architectures and sharding.
- Demonstrated engineering expertise, including gathering requirements, problem-solving, and making recommendations.
- Preferred: Familiarity with security frameworks, attack vectors, botnets, and impact analysis.
What we offer:
- Fully remote position with flexible working hours.
- An inspiring team of colleagues spread all over the world.
- Pleasant, modern development and deployment workflows: ship early, ship often.
- High impact: lots of users, happy customers, high growth, and cutting edge R&D.
- Flat organization, direct interaction with customer teams.
We celebrate equality of opportunity and are committed to creating an inclusive environment for all team members. Join us as we transform cybersecurity, user privacy, and machine learning online!
Please note that all positions require pre-employment screening, including third-party verification of work history, education, and identity, as well as a final in-person interview and identity verification step, which will be conducted in your country of residence.
Originally posted on Himalayas