远程工作雷达

高级软件工程师 – BMaaS 与数据中心网络

Senior Software Engineer – BMaaS & Datacenter Networking

开发工程限定地区(需当地身份)
公司Mirantis
薪资未公开
工作地点United States
地域资格限定地区(需当地身份)
时区要求日间重叠约 9 小时,基本正常作息
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 United States 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

关于团队
我们的工程团队为AI新云环境构建下一代、原生云的裸金属编排平台。我们采用T型工程模型:每位工程师在其主领域拥有深厚的专业知识,同时每个团队成员在相邻系统中保持实际的熟练度。这种共享基础使我们能够进行严谨的设计评审、有效的跨领域代码评审和可靠的值班支持。我们的开发以开源项目为基础,参与开源项目是我们的主要工作之一。
我们使用Rust和Go开发核心控制平面软件,利用Kubernetes operators,并推动以OpenTelemetry为核心的运维文化。
职位概述
我们正在寻找一位资深软件工程师,具备扎实的软件架构原则,并在数据中心网络和DPU架构方面有深入的技术领域知识。符合我们的T型工程模型,您将推动为AI训练和推理工作负载提供支持的高密度、高吞吐量基础设施的控制平面开发。熟悉自动化裸金属配置和生命周期管理是一项重要优势。
核心职责

  • 控制平面开发:使用Rust和Go设计、构建和维护生产级控制平面微服务、自定义Kubernetes operator和强大的协调引擎。
  • 架构与API设计:编写清晰的功能(FR)和非功能需求(NFR)、架构规范、系统时序图以及干净的gRPC/Protobuf和REST API模式。
  • 网络集成:为现代网络操作系统(SONiC、NVUE、Cumulus)和DPU硬件平台开发定制集成模块和集成方案。
  • 可观测性与诊断:使用OpenTelemetry(OTel)追踪和指标对服务进行端到端监控,对多语言分布式系统进行系统的根本原因分析。
  • 质量与工程卓越:参与针对Rust、Go和SQL代码库的严格、评审通过的拉取请求流程,同时通过模拟驱动测试确保高测试覆盖率。
  • 技术要求
  • 语言能力:精通Rust(Tokio异步运行时、Tonic、Axum、sqlx),并具有Go的高水平技能。
  • 数据与API:高级SQL/PostgreSQL熟练度,gRPC/Protobuf契约设计和模式演进。
  • 并发与可靠性:对异步并发模型、无锁模式和分布式状态处理有扎实的背景。
查看英文原文

About the Team
Our engineering team builds next-generation, cloud-native bare-metal orchestration platforms for AI neo-cloud environments. We operate on a T-shaped engineering model: while each engineer brings deep expertise in a primary domain, every team member maintains practical fluency across neighboring systems. This shared foundation enables rigorous design reviews, effective cross-domain code reviews, and dependable on-call coverage. Our development is fundamentally based on open-source projects, and contributing to them is a major part of our work.
We develop primary control plane software in Rust and Go, leverage Kubernetes operators, and foster an OpenTelemetry-first operational culture.
Role Overview
We are seeking a Senior Software Engineer who combines strong software architecture principles with deep technical domain expertise in Datacenter Networking and DPU architectures. Fitting into our T-shaped engineering model, you will drive the development of control planes powering high-density, high-throughput infrastructure for AI training and inference workloads. Familiarity with automated bare-metal provisioning and lifecycle management is a strong asset.
Core Responsibilities

  • Control Plane Development: Design, build, and maintain production-grade control plane microservices, custom Kubernetes operators, and robust reconciliation engines in Rust and Go.
  • Architecture & API Design: Author clear Functional (FR) and Non-Functional Requirements (NFR), architectural specs, system sequence diagrams, and clean gRPC/Protobuf and REST API schemas.
  • Network Integration: Develop custom integration modules and integrations for modern network operating systems (SONiC, NVUE, Cumulus) and DPU hardware platforms.
  • Observability & Diagnostics: Instrument services end-to-end using OpenTelemetry (OTel) traces and metrics, performing systematic root-cause analysis across polyglot distributed systems.
  • Quality & Engineering Excellence: Participate in rigorous, review-gated pull request workflows across Rust, Go, and SQL codebases while ensuring high test coverage via mock-driven testing.

Technical Qualifications

  • Language Proficiency: Primary mastery of Rust (Tokio async runtime, Tonic, Axum, sqlx) and strong proficiency in Go.
  • Data & APIs: Advanced SQL / PostgreSQL fluency, gRPC/Protobuf contract design, and schema evolution.
  • Concurrency & Reliability: Strong background in async concurrency models, lock-free patterns, distributed state handling, and mock-driven testing discipline.
  • Datacenter Protocols & OS: Expertise in BGP, MP-BGP, EVPN, VXLAN, L3VNI, route targets, and route server design. Hands-on experience with SONiC, Cumulus Linux, and NVUE (featuring a first-class NVUE client).
  • DPU & Fabric Ecosystem: Deep knowledge of NVIDIA DOCA, Host-Based Networking (HBN), BlueField DPU architectures, and the DPF (DOCA Platform Framework) operator model (BFB, DPUSet, DPUNode CRDs).
  • High-Performance Interconnects: Understanding of InfiniBand fabrics, NVLink / NMX-M partitioning, and RoCEv2.
  • Linux Kernel Networking: Advanced grasp of Linux netlink, network namespaces, routing tables, and internal DHCP/DNS service implementations (the project ships its own services).
  • Cloud & Kubernetes Networking: Familiarity with CNI plugins (Calico, Cilium). OVS and DPDK experience is a plus.
  • Provisioning & Boot Infrastructure: Expertise in PXE/iPXE, UEFI, Secure Boot, measured boot, and TPM attestation.
  • Hardware Management: Proficiency in Redfish, IPMI, and BMC abstractions across heterogeneous hardware platforms (Dell, Lenovo, NVIDIA reference hardware).
  • Linux Systems Internals: In-depth knowledge of boot chains, systemd, initramfs, BIOS configuration matrices.
  • Identity & Security: Practical experience with PKI, X.509 certificates, TLS, SPIFFE/SVID, Vault, KMS, Keycloak (OAuth2/JWT), and RBAC.
  • Virtualization: Experience with KubeVirt / KVM.
  • Kubernetes Control Planes: Experience writing custom controllers/operators using kube-rs or controller-runtime with CRD-driven reconciliation loop patterns.

Universal Requirements

  • Testing Discipline: Commitment to test-driven design, writing highly testable code with mock interfaces for external hardware and network dependencies.
  • Collaborative Architecture: Proven track record of writing crisp architectural specs and engaging in constructive, cross-functional design reviews.

Educational & Experience Baseline

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or equivalent practical experience.
  • 5+ years of software engineering experience in cloud-native platforms, systems programming, networking, or infrastructure automation.
  • Demonstrated participation or maintainership in open-source systems projects is a plus.

What does Mirantis offer you?

  • Work with an established Silicon Valley leader in the cloud infrastructure industry;
  • Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
  • Be a part of cutting-edge, open-source innovation;
  • Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
  • Professional development and training;
  • Attend conferences and working groups;
  • Company outings, happy hours, hackathons, and tech talks;
  • Receive a competitive compensation package with a strong benefits plan.

We are a Leader for Container Management in G2 (#2 after AWS)!
Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位