远程工作雷达

高级基础设施与 DevOps 工程师

Senior Infrastructure & DevOps Engineer

开发工程职能支持未标注地域
公司XTIUM
薪资未公开
工作地点Pakistan
地域资格未标注地域
时区要求日间重叠约 6 小时,基本正常作息
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →

XTIUM全球团队由一群多元且才华横溢的专业人士组成,大家有着共同的目标:追求卓越和持续改进。我们致力于迎接挑战,保持沟通畅通并协作共进。我们对工作负责,注重学习与成长,并对同事和客户保持问责。我们共同努力突破界限,产生影响,并激励彼此发挥全部潜力。
职位描述:
1. 基础设施与操作系统管理

  • 管理和维护生产及测试环境中的基于Linux的系统(Rocky Linux、CentOS、RHEL)。
  • 按照行业最佳实践执行内核补丁、安全更新和系统加固。
  • 排查和调试复杂的操作系统级问题(性能、内存、I/O、内核崩溃)。
  • 在适用情况下管理和支持Windows Server环境。

2. 安全与合规

  • 使用Nessus、Qualys或OpenVAS等工具进行定期安全扫描和漏洞评估。
  • 执行代码安全扫描(SAST/DAST)并将其集成到CI/CD流程中。
  • 配置和维护LDAP认证以实现集中式访问管理。
  • 确保所有系统符合内部安全策略和外部监管标准。
  • 实现和管理WAF(Web应用防火墙)配置,以防范OWASP Top 10威胁。

3. 自动化与配置管理

  • 编写和维护Ansible剧本,用于自动化部署、补丁和配置管理。
  • 管理GitHub Enterprise仓库、分支策略和CI/CD流程(优先使用GitHub Actions)。
  • 使用Python、Shell或Go开发自动化脚本,减少重复劳动并提高运营效率。

4. 应用与Web服务器管理

  • 配置和维护Apache和Nginx Web服务器,包括性能调优和反向代理设置。
  • 排查Web服务器日志、SSL/TLS问题和负载均衡配置。

5. 容器编排与云原生技术

  • 在Kubernetes集群(EKS、AKS或自建)上部署、管理和扩展应用。
  • 构建和维护具有安全基础层和最小攻击面的Docker镜像。
  • 监控集群健康状况、资源使用情况,并实施自动扩展策略。

6. 监控与事件管理

  • 部署和维护监控工具(Prometheus、Grafana、Datadog或ELK堆栈)。
  • 定义SLIs、SLOs和错误率指标。
查看英文原文

The XTIUM global team is made up of a group of diverse and talented professionals who are all driven by the same goal: excellence and continuous improvement. We are all about embracing challenges, keeping the lines of communication open and working together. We take ownership of our work, focus on learning and growing and hold ourselves accountable to our colleagues and customers. Together, we strive to push boundaries, make an impact and inspire each other to reach our full potential.
Job Description:
1. Infrastructure & OS Management

  • Manage and maintain Linux-based systems (Rocky Linux, CentOS, RHEL) in production and staging environments.
  • Perform kernel patching, security updates, and system hardening following industry best practices.
  • Troubleshoot and debug complex OS-level issues (performance, memory, I/O, kernel panics).
  • Manage and support Windows Server environments where applicable.

2. Security & Compliance

  • Perform regular security scanning and vulnerability assessments using tools like Nessus, Qualys, or OpenVAS.
  • Conduct code security scans (SAST/DAST) and integrate them into CI/CD pipelines.
  • Configure and maintain LDAP authentication for centralized access management.
  • Ensure all systems comply with internal security policies and external regulatory standards.
  • Implement and manage WAF (Web Application Firewall) configurations to protect against OWASP Top 10 threats.

3. Automation & Configuration Management

  • Write and maintain Ansible playbooks for automated provisioning, patching, and configuration management.
  • Manage GitHub Enterprise repositories, branching strategies, and CI/CD workflows (GitHub Actions preferred).
  • Develop automation scripts using Python, Shell, or Go to reduce toil and improve operational efficiency.

4. Application & Web Server Management

  • Configure and maintain Apache and Nginx web servers, including performance tuning and reverse proxy setup.
  • Troubleshoot web server logs, SSL/TLS issues, and load balancing configurations.

5. Container Orchestration & Cloud-Native Technologies

  • Deploy, manage, and scale applications on Kubernetes clusters (EKS, AKS, or on-prem).
  • Build and maintain Docker images with secure base layers and minimal attack surfaces.
  • Monitor cluster health, resource usage, and implement auto-scaling policies.

6. Monitoring & Incident Management

  • Deploy and maintain monitoring tools (Prometheus, Grafana, Datadog, or ELK stack).
  • Define SLIs, SLOs, and error budgets; drive blameless post-mortems.
  • Participate in on-call rotations and lead incident response efforts.

What Qualifies You

  • 8+ years in SRE, DevOps, or Systems Engineering roles.
  • Experience with Terraform or other IaC tools.
  • Familiarity with service mesh (Istio/Linkerd) and CNI plugins (Calico/Cilium) or similar.
  • Understanding of zero-trust networking and SSO integrations.
  • Certifications: RHCE, CKA, or CISSP are a plus.

OS: Expert in Rocky Linux, CentOS, RHEL (6+ years); Windows Server experience
Security: WAF, vulnerability scanning, LDAP, OS hardening, kernel patching
Automation: Ansible (playbooks, roles, AWX/Tower), Python/Shell scripting
CI/CD/SCM: GitHub Enterprise administration, GitHub Actions, branch protection rules
Containers: Docker (build/security scanning), Kubernetes (deployment, networking, storage)
Web Servers: Apache, Nginx (configuration, SSL, reverse proxy)
Monitoring: Prometheus, Grafana, ELK, or commercial tools (Datadog/New Relic)
Code Security: SAST/DAST tools (SonarQube, Snyk, Checkmarx)
Soft Skills & Cultural Fit

  • Strong communication and documentation skills.
  • Ability to mentor junior engineers and conduct technical interviews.
  • Collaborative mindset with DevOps/DevSecOps philosophy.
  • Calm under pressure during incidents; drive for continuous improvement.

XTIUM is an equal opportunity employer.
RemoteOriginally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

← 返回全部职位