系统工程师/DevOps – 高级
System Engineer/ DevOps – Senior
Overview:
该职位需要专注于设计、扩展和保障分布式环境,同时将基础设施视为内部产品。该职位是软件开发与核心基础设施之间的重要桥梁。
该职位重点是构建抽象的内部工具,设定工程守则,并作为软件开发团队的高级技术顾问。该职位需要在优化分布式系统、设计可扩展的云架构以及指导工程团队遵循云原生最佳实践方面具备深厚的专业知识。
About Product:
SOFTSWISS Game Aggregator
一种快速且成本效益高的解决方案,可让您轻松集成和管理赌场游戏内容
了解更多信息
Purpose of the role:
您将设计、构建和优化可扩展的基础设施和部署流水线,以确保我们平台上的系统可靠性、性能和安全性。您的工作将通过实现高效的 CI/CD 流程、提升可观测性以及支持开发团队更快地发布高质量功能来影响服务的正常运行时间和交付速度。
Key responsibilities:
- 使用可重用的 Helm 模板设计和维护可扩展的多租户 Kubernetes 平台,以实现开发者自助服务。
- 负责流水线的可靠性、成本和速度。嵌入自动化发布门禁、安全的密钥架构和软件漏洞扫描(SCA)。
- 构建自定义工具以扩展平台功能并标准化共享环境。
- 维护分布式追踪、日志和指标系统。指导开发团队进行服务仪器化并建立 SLIs/SLOs/SLAs。
- 为开发团队提供系统架构、资源优化和现代流量路由模式方面的咨询。
Required Experience:
- 5年以上在 DevOps、SRE 或平台工程岗位中支持高负载分布式系统的经验。
- 对多云模式、多区域架构和云成本优化有深入理解。
- 具备生产级 Kubernetes 管理经验,包括高级调度策略、网络隔离和复杂 Helm 模板。
- 深入的 Linux 管理和跨堆栈故障排查能力。有配置 L4/L7 流量层(负载均衡器、Ingress、API 网关、服务网格)的经验。
- 精通声明式 IaC 并熟练掌握 Python、Go 或 Bash。
- 能够独立收集需求、设定交付目标并构建可维护的解决方案。
- 中级或高级水平的英语沟通能力。
查看英文原文
Overview:
The role requires focusing on designing, scaling, and securing distributed environments while treating infrastructure as an internal product. This position serves as a critical bridge between software development and the core infrastructure.
This role focuses on building abstract internal tooling, setting engineering guardrails, and serving as a high-level technical advisor to software development teams. The position requires deep expertise in optimizing distributed systems, designing scalable cloud architectures, and mentoring engineering teams on cloud-native best practices.
About Product:
SOFTSWISS Game Aggregator
A fast and cost-effective solution that allows you to integrate and manage casino gaming content easily
Learn More
Purpose of the role:
You’ll design, build, and optimise scalable infrastructure and deployment pipelines to ensure high system reliability, performance, and security across our platforms. Your work will impact service uptime and delivery speed by enabling efficient CI/CD processes, improving observability, and supporting development teams in releasing high-quality features faster.
Key responsibilities:
- Design and maintain scalable, multi-tenant Kubernetes platforms using reusable Helm templates to enable developer self-service.
- Own pipeline reliability, cost, and speed. Embed automated release gates, secure secret architectures, and software vulnerability scanning (SCA).
- Build custom tools to extend platform capabilities and standardize shared environments.
- Maintain distributed tracing, logging, and metrics systems. Guide development teams on service instrumentation and establishing SLIs/SLOs/SLAs.
- Consult development teams on system architecture, resource optimization, and modern traffic routing patterns.
Required Experience:
- 5+ years in a DevOps, SRE, or Platform Engineering role supporting high-load distributed systems.
- Strong grasp of multi-cloud patterns, multi-region architectures, and cloud cost optimization.
- Production-grade Kubernetes administration, including advanced scheduling policies, network isolation, and complex Helm templating.
- Deep Linux administration and cross-stack troubleshooting. Experience configuring L4/L7 traffic layers (Load Balancers, Ingress, API Gateways, Service Meshes).
- Mastery of declarative IaC and proficiency in Python, Go, or Bash.
- Ability to independently gather requirements, set delivery goals, and build maintainable solutions.
- Intermediate or higher English and Russian (B1+) proficiency.
Nice to have:
· Practical experience scaling and maintaining infrastructure for AI-powered applications, LLM frameworks, or MLOps pipelines.
Our Benefits:
- Private health insurance
- Sports benefits
- Comprehensive Mental Health Program
- Free English lessons (online)
- Local language courses
- Paid time off
- Maternity leave support
- Referral program rewards
- Upskilling, internal workshops, and participation in professional conferences and corporate events
Originally posted on Himalayas