资深 DevOps 工程师
Sr DevOps Engineer
“为什么微服务得到了晋升?
因为它在压力下依然能够扩展。”
我们正在为一家备受信赖的客户招聘一名经验丰富的DevOps工程师,这是一家专注于AI驱动的品牌安全和上下文智能解决方案的创新技术公司。该职位专注于构建和维护支持其先进数据分析和内容测量平台的稳健基础设施基础。
理想的候选人将设计、保障并优化使可扩展、高性能应用程序得以运行的操作系统。对于热衷于自动化、系统负责制以及从零开始构建弹性基础设施解决方案的人来说,这是一个关键角色。
主要职责
- 设计、实施并维护用于高吞吐量、低延迟应用的安全、可扩展云基础设施(主要为AWS,也具备多云环境的灵活性)
- 开发和增强CI/CD流水线,以实现高效、可靠且一致的部署流程(GitHub Actions及类似工具)
- 持续识别并实施整个产品生态系统的基础设施改进
- 架构并支持使用AWS Lambda、ECS Fargate和事件驱动架构组件的无服务器解决方案
- 建立全面的监控、警报和日志框架,以确保系统可靠性与可见性
- 使用CloudFormation、Terraform及相关工具管理基础设施即代码的实现
- 与开发团队合作,确定基础设施需求、部署方法和运营工具集
- 管理密钥管理和身份系统,使用AWS IAM及其他平台
- 维护符合安全和隐私标准的合规性,包括访问控制、加密协议和审计机制
- 跨所有服务层解决生产事故,注重快速响应和彻底的事后分析
- 创建和维护自动备份、灾难恢复和故障转移系统
- 研究并采用新兴的DevOps方法和技术,以提升平台性能和可靠性
所需资格
- 10年以上软件工程背景,其中6年以上专注于DevOps、站点可靠性工程或基础设施工程
- 有管理生产环境并承担全面基础设施责任的实际经验
- 丰富的AWS专业知识,包括计算、存储、网络和
查看英文原文
“Why did the microservice get promoted?
Because it scaled under pressure.”
We are recruiting a seasoned DevOps Engineer for one of our valued clients, an innovative technology company specializing in AI-powered brand safety and contextual intelligence solutions. This role focuses on building and maintaining the robust infrastructure foundation that supports their advanced data analytics and content measurement platforms.
The ideal candidate will design, secure, and optimize the operational systems that enable scalable, high-performance applications. This is a pivotal role for someone passionate about automation, system ownership, and creating resilient infrastructure solutions from the ground up.
Key Responsibilities
- Design, implement, and maintain secure, scalable cloud infrastructure for high-throughput, low-latency applications (mainly AWS, with flexibility for multi-cloud environments)
- Develop and enhance CI/CD pipelines for efficient, reliable, and consistent deployment processes (GitHub Actions and similar tools)
- Continuously identify and implement infrastructure improvements across the entire product ecosystem
- Architect and support serverless solutions using AWS Lambda, ECS Fargate, and event-driven architecture components
- Establish comprehensive monitoring, alerting, and logging frameworks to ensure system reliability and visibility
- Manage infrastructure-as-code implementations using CloudFormation, Terraform, and related tools
- Partner with development teams to determine infrastructure requirements, deployment approaches, and operational toolsets
- Oversee secrets management and identity systems utilizing AWS IAM and similar platforms
- Maintain compliance with security and privacy standards, including access controls, encryption protocols, and audit mechanisms
- Resolve production incidents across all service layers with a focus on rapid response and thorough post-incident analysis
- Create and maintain automated backup, disaster recovery, and failover systems
- Research and adopt emerging DevOps methodologies and technologies to enhance platform performance and reliability
Required Qualifications
- 10+ years of software engineering background with 6+ years focused on DevOps, Site Reliability Engineering, or Infrastructure Engineering
- Demonstrated experience managing production environments with comprehensive infrastructure responsibilities
- Extensive AWS expertise, including compute, storage, networking, and identity management services
- Practical experience with serverless technologies (AWS Lambda, Step Functions, EventBridge, API Gateway, ECS Fargate)
- Advanced skills in Docker and container orchestration platforms (Kubernetes, ECS, GKE)
- Proficiency with CI/CD platforms such as GitHub Actions, CircleCI, ArgoCD, or Jenkins
- Strong scripting capabilities in Bash, Python, or Go for automation and tooling development
- Experience with observability solutions (Datadog, Prometheus, Grafana, ELK stack)
- Solid understanding of network design, security frameworks, and zero-trust access architectures
- Knowledge of secrets management systems and infrastructure-level access policy enforcement
- Exceptional troubleshooting and root cause analysis abilities
- Strong collaborative and communication skills across diverse technical and business teams
- Continuous improvement mindset with focus on automation, optimization, and security enhancement
This position offers the opportunity to join an early-stage team with proven success in developing cutting-edge contextual intelligence technology, working alongside talented engineers and product leaders who prioritize innovation and measurable client impact.
What are you waiting for? Fill out the form below and apply!