网络基础设施软件工程师
Software Engineer in Network Infrastructure
关于Nebius:
Nebius正在引领全球AI经济的云基础设施新时代。我们正在构建一个全栈AI云平台,支持开发者和企业从数据和模型训练到生产部署,而无需承担构建大型内部AI/ML基础设施的成本和复杂性。
由工程师打造,为工程师而生。从大规模GPU编排到推理优化,我们在计算、存储、网络和应用AI领域都负责解决难题。
在纳斯达克上市(股票代码:NBIS),总部位于阿姆斯特丹,我们在欧洲、英国、北美和以色列设有研发中心,拥有全球业务布局。我们的团队超过1500人,包括数百名在硬件、软件和AI研发方面有深厚专业知识的工程师。
职位描述
我们正在寻找一名软件工程师,开发使网络运维安全、可扩展且枯燥的软件——即使我们在快速推出新数据中心并扩展业务。这不是一份“编写脚本配置”的工作:你将构建位于网络核心(交换机/端口/VLAN、流量处理器)和上层云平台之间的工具和系统,使用开源技术,并在需要时构建缺失的部分。
你的职责将包括:
- 构建和维护自动化网络生命周期的服务和工具:初始配置、日常变更、漂移检测和操作验证
- 使网络变更安全透明:CI/CD流程、差异/评审工具、分阶段发布/回滚、审计日志和防护机制
- 开发可跨多个站点扩展的可观测系统:遥测管道、信号质量以及缩短事件调查的工具
- 填补网络与平台之间的“最后一公里”差距:集成真实数据源、暴露API,并围绕其构建可靠的自动化
- 与网络工程师和SRE紧密合作,将实际的运维痛点转化为高质量的工具
我们期望你具备:
- 5年以上专业软件工程经验(或同等实践经验)
- 强大的编码能力和主人翁意识:你能交付并运营可靠的服务
- 精通Go语言或愿意切换;Python也欢迎(其他语言在开源调试/修复中可能有用)
- 你不必是网络专家,但我们期望你对基础设施和/或网络有兴趣
如果你有以下背景,将是一个加分项:
- 网络相关背景(如曾是网络工程师,持有CCNP/ed认证等)
查看英文原文
About Nebius:
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.
Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
The Role
We’re looking for a Software Engineer to build software that makes network operations safe, scalable, and boring — even as we launch new data centers and expand fast. This is not a “write scripts for configs” role: you’ll build the tooling and services that sit between the network core (switches/ports/VLANs, traffic processors) and the cloud platform on top, using open source where it fits and building the missing pieces where it doesn’t.
Your responsibilities will include:
- Build and maintain services and tooling that automate the network lifecycle: day-0 provisioning, day-N changes, drift detection, and operational verification
- Make network changes safe and transparent: CI/CD workflows, diff/review tooling, staged rollouts/rollbacks, audit trails, and guardrails
- Develop observability systems that scale across many sites: telemetry pipelines, signal quality, and tooling that shortens incident investigations
- Close “last mile” gaps between the network and the platform: integrate source-of-truth data, expose APIs, and build reliable automation around it
- Collaborate closely with network engineers and SREs to turn real operational pain into product-quality tooling
We expect you to have:
- 5+ years of professional software engineering experience (or equivalent practical background)
- Strong coding skills and ownership mindset: you can ship and operate reliable services
- Proficiency in Go or readiness to switch; Python is also welcome (other languages can be useful for OSS debugging/fixes)
- You don’t have to be Network expert but we expect you to have interest in infrastructure and/or network
It will be an added bonus if you have:
- Background in networking (ex-network engineer, CCNP/education, DC networking exposure) or strong interest and proven ability to learn fast
- Experience building automation/infra tooling: CI/CD, IaC, testing/staging environments, or “network-as-code” style workflows
- Low-level networking / datapath experience: eBPF/XDP, DPDK, kernel networking, traffic processing systems
- Experience designing high-load services and observability platforms (metrics/logs/traces, alerting, regression detection)
- Contributions to open source or experience extending/debugging OSS components in production environments
We expect Engineers to:
- Manage large-scale projects involving multiple stakeholders
- Break down complex tasks and guide both their own work and that of more junior colleagues
- Be experts in specific technologies and write high-quality code that can serve as a reference
- Assess task priority and focus on high-impact work, avoiding low-value efforts
- Have strong architectural thinking and contribute to system design
- Be involved in hiring and actively contribute to interviews
- Be willing to share knowledge and mentor others
Benefits & Perks:
- Competitive compensation
- Career growth and learning opportunities
- Flexibility and ownership
- Collaborative and innovative culture
- Opportunity to work on impactful AI projects
- International environment and talented teams
What's it like to work at Nebius:
Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI
Equal Opportunity Statement:
Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.
Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire.
If you need accommodations during the application process, please let us know.