远程工作雷达

AWS 云工程师 III

AWS Cloud Engineer III

开发工程限定地区(需当地身份)
公司Rackspace US, Inc.
薪资未公开
工作地点India
地域资格限定地区(需当地身份)
时区要求日间重叠约 6 小时,基本正常作息
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 India 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

职位概述
L3 DevOps/云工程师将作为高级技术资源,负责主导复杂的基础设施和DevOps活动,推动自动化,定义技术标准,并提升平台可靠性。
该职位需要在AWS、Terraform、AWX/Ansible、ArgoCD、GitOps、Kubernetes、CI/CD和灾难恢复方面具备扎实的实操经验。
主要职责

  • 作为复杂云、DevOps、Kubernetes、基础设施和部署问题的L3级支持人员。
  • 在生产和非生产环境中主导主要技术活动。
  • 负责复杂基础设施变更、升级、迁移和平台改进。
  • 主导并协调灾难恢复(DR)活动,包括:
  • DR计划
  • 恢复测试
  • 故障转移/故障恢复活动
  • 应用和基础设施恢复
  • DR验证
  • 操作手册的准备和改进
  • 使用Terraform设计和实现基础设施。
  • 开发和维护可重用的Terraform模块。
  • 审查L1/L2工程师实施的Terraform代码和基础设施变更。
  • 定义Terraform编码标准、仓库结构和实现最佳实践。
  • 在AWX/Ansible方面有扎实的实操经验。
  • 开发可重用的Ansible角色、剧本和自动化工作流。
  • 使用AWX自动化:
  • 服务器配置
  • 补丁管理
  • 应用部署
  • 基础设施操作
  • 重复性的日常运营活动
  • 识别手动操作流程并将其转换为自动化工作流。
  • 设计和维护基于ArgoCD的部署。
  • 实施并支持Kubernetes和应用部署的GitOps实践。
  • 定义GitOps仓库结构和部署标准。
  • 管理和排查ArgoCD:
  • 应用程序
  • 同步问题
  • 配置漂移
  • 部署失败
  • 环境推广
  • 回滚活动
  • 设计和维护运行在Amazon EKS上的Kubernetes工作负载。
  • 开发和维护可重用的Helm图表。
  • 定义Helm值、模板和环境特定配置的标准。
  • 设计和改进CI/CD部署流程。
  • 支持并改进Jenkins和其他CI/CD自动化。
  • 定义部署策略和操作标准。
  • 在云和DevOps平台上主导自动化项目。
  • 使用以下工具开发自动化:
  • Terraform
  • Ansible/AWX
  • Bash
  • Python
  • CI/CD流水线
  • 定义并执行技术标准
查看英文原文

Role Overview
The L3 DevOps / Cloud Engineer will act as a senior technical resource responsible for leading complex infrastructure and DevOps activities, driving automation, defining technical standards, and improving platform reliability.
The role requires strong hands-on expertise in AWS, Terraform, AWX/Ansible, ArgoCD, GitOps, Kubernetes, CI/CD, and Disaster Recovery.
Key Responsibilities

  • Act as the L3 escalation point for complex Cloud, DevOps, Kubernetes, infrastructure, and deployment issues.
  • Lead major technical activities across production and non-production environments.
  • Take ownership of complex infrastructure changes, upgrades, migrations, and platform improvements.
  • Lead and coordinate Disaster Recovery (DR) activities, including:
  • DR planning
  • Restore testing
  • Failover/failback activities
  • Application and infrastructure recovery
  • DR validation
  • Runbook preparation and improvement
  • Design and implement infrastructure using Terraform.
  • Develop and maintain reusable Terraform modules.
  • Review Terraform code and infrastructure changes implemented by L1/L2 engineers.
  • Define Terraform coding standards, repository structure, and implementation best practices.
  • Strong hands-on experience with AWX / Ansible.
  • Develop reusable Ansible roles, playbooks, and automation workflows.
  • Use AWX to automate:
  • Server configuration
  • Patching
  • Application deployment
  • Infrastructure operations
  • Repetitive BAU activities
  • Identify manual operational activities and convert them into automated workflows.
  • Design and maintain ArgoCD-based deployments.
  • Implement and support GitOps practices for Kubernetes and application deployments.
  • Define GitOps repository structures and deployment standards.
  • Manage and troubleshoot ArgoCD:
  • Applications
  • Sync issues
  • Configuration drift
  • Deployment failures
  • Environment promotion
  • Rollback activities
  • Design and maintain Kubernetes workloads running on Amazon EKS.
  • Develop and maintain reusable Helm Charts.
  • Define standards for Helm values, templates, and environment-specific configurations.
  • Design and improve CI/CD deployment processes.
  • Support and improve Jenkins and other CI/CD automation.
  • Define deployment strategies and operational standards.
  • Lead automation initiatives across Cloud and DevOps platforms.
  • Develop automation using:
  • Terraform
  • Ansible / AWX
  • Bash
  • Python
  • CI/CD pipelines
  • Define and enforce technical standards for:
  • Infrastructure as Code
  • Terraform
  • GitOps
  • Kubernetes
  • Helm
  • CI/CD
  • Automation
  • Cloud operations
  • Patching
  • Deployment processes
  • Establish reusable templates, modules, pipelines, and automation frameworks.
  • Perform technical reviews for changes implemented by L1 and L2 engineers.
  • Lead complex production incidents and perform Root Cause Analysis.
  • Identify recurring issues and implement permanent automated solutions.
  • Improve monitoring, logging, alerting, and operational reliability.
  • Participate in architecture and technical design discussions.
  • Lead production deployments, infrastructure upgrades, patching, and maintenance activities.
  • Mentor and provide technical guidance to L1 and L2 engineers.
  • Create and maintain:
  • SOPs
  • Runbooks
  • Technical standards
  • Architecture documentation
  • DR documentation
  • Operational procedures

Experience
8–12 years of relevant experience in cloud engineering, infrastructure, DevOps, or a related field is required.
Core Technical Skills
Strong hands-on expertise in:

  • AWS
  • Amazon EKS
  • Kubernetes
  • Terraform
  • Terragrunt
  • AWX
  • Ansible
  • ArgoCD
  • GitOps
  • Helm
  • Docker
  • Jenkins
  • CI/CD
  • Git
  • Linux
  • Windows
  • Bash / Python
  • SSL/TLS
  • Infrastructure Automation
  • Disaster Recovery
  • Monitoring & Logging
  • Production Troubleshooting
  • Root Cause Analysis

L3 Expectations
An L3 Engineer should be capable of:

  • Leading technical activities independently
  • Owning complex Cloud and DevOps changes
  • Designing and implementing automation
  • Leading Disaster Recovery activities
  • Implementing GitOps using ArgoCD
  • Building automation using AWX/Ansible
  • Developing and reviewing Terraform code
  • Defining technical standards and best practices
  • Driving operational improvements
  • Mentoring L1/L2 engineers
  • Handling complex production incidents and RCA

The L3 engineer should move beyond routine operational support and focus on engineering, automation, standardization, reliability, and technical leadership.
About Rackspace Technology
We are the multicloud solutions experts. We combine our expertise with the world’s leading technologies — across applications, data and security — to deliver end-to-end solutions. We have a proven record of advising customers based on their business challenges, designing solutions that scale, building and managing those solutions, and optimizing returns into the future. Named a best place to work, year after year according to Fortune, Forbes and Glassdoor, we attract and develop world-class talent. Join us on our mission to embrace technology, empower customers and deliver the future.
More on Rackspace Technology
Though we’re all different, Rackers thrive through our connection to a central goal: to be a valued member of a winning team on an inspiring mission. We bring our whole selves to work every day. And we embrace the notion that unique perspectives fuel innovation and enable us to best serve our customers and communities around the globe. We welcome you to apply today and want you to know that we are committed to offering equal employment opportunity without regard to age, color, disability, gender reassignment or identity or expression, genetic information, marital or civil partner status, pregnancy or maternity status, military or veteran status, nationality, ethnic or national origin, race, religion or belief, sexual orientation, or any legally protected characteristic. If you have a disability or special need that requires accommodation, please let us know.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

AWS云工程师IV

Rackspace US, Inc.IndiaFull Time今天
开发工程限定地区(需当地身份)

云工程师 II

Rackspace US, Inc.IndiaFull Time今天
开发工程职能支持限定地区(需当地身份)

应用集成顾问 V

Rackspace US, Inc.IndiaFull Time今天
开发工程限定地区(需当地身份)

医疗保健部门营销经理

Rackspace US, Inc.United States$99,704 - $146,270.3/年Full Time今天
市场运营职能支持限定地区(需当地身份)

← 返回全部职位