高级资深系统工程师 - 性能工程师
Senior Staff Systems Engineer - Performance Engineer
Nu 是拉丁美洲领先的数字银行,为巴西、墨西哥和哥伦比亚的 1.4 亿客户提供服务。公司通过利用数据和专有技术开发创新产品和服务,引领行业变革。
以“对抗复杂性,赋能人们”为使命,Nu 为客户完整的金融旅程提供服务,通过负责任的贷款和透明度促进金融包容性和进步。公司采用高效且可扩展的商业模式,结合低成本服务与不断增长的回报。
Nu 的影响力已获得多项奖项的认可,包括《时代》杂志 100 家最具影响力公司、《快公司》最具创新力公司以及《福布斯》全球最佳银行。
访问我们的机构页面 https://www.nu.com/2026-en
职位简介
系统性能团队是计算小组的一部分,包含两个不同的工作流——编排和系统性能——每个工作流都有自己的方法和挑战,负责管理并提升大多数 Nubank 工作负载运行的基础架构。
性能团队专注于构建深入的诊断工具,并进行高层次分析,以降低延迟、基础设施成本并提高服务效率。
您将负责领导复杂的性能调查,识别系统性瓶颈,并推动全球最大的基于 JVM 的微服务架构的效率提升。
我们的核心原则和行为包括责任感、简洁性、真实性优先、团队合作以及注重质量而非数量。在正常的工作日中,您将与关键基础设施层互动,从 Linux 内核和 JVM 内部结构到云级编排。
您将负责:
- 领导深度调查:进行高层次性能分析,识别并解决我们全球基于 JVM 的微服务架构中的系统性瓶颈。
- 优化资源效率:通过调整 JVM 参数、垃圾回收(ZGC、G1)和内存管理(堆内存和非堆内存)来推动降低基础设施成本和延迟的举措。
- 构建诊断工具:使用 eBPF、JFR 和火焰图开发并实现高级可观测性工具,提供对内核和运行时行为的实时洞察。
- 内核与运行时对齐:弥合 Linux 内核和 JVM 之间的差距,优化线程调度(
查看英文原文
ABOUT NU
Nu is the leading digital bank in Latin America, serving 140 million customers across Brazil, Mexico, and Colombia. The company has been leading an industry transformation by leveraging data and proprietary technology to develop innovative products and services.
Guided by its mission to fight complexity and empower people, Nu caters to customers’ complete financial journey, promoting financial access and advancement with responsible lending and transparency. The company is powered by an efficient and scalable business model that combines low cost to serve with growing returns.
Nu’s impact has been recognized in multiple awards, including Time 100 Most Influential Companies, Fast Company’s Most Innovative Companies, and Forbes World’s Best Banks.
Visit our Institutional Page https://www.nu.com/2026-en
About the Role
The Systems Performance team is part of the Computing Squad consists of two distinct workstreams —Orchestration and System Performance —each with its own approach and challenges on managing and improving the foundational infrastructure where the majority of the Nubank's workloads runs.
The Performance team is focused on building deep diagnostic tools and performing high-level analysis to reduce latency, infrastructure costs and increase services efficiency.
You will be responsible for leading complex performance investigations, identifying systemic bottlenecks, and driving efficiency across one of the largest JVM-based microservice architectures in the world.
Our core principles and behaviors include ownership, simplicity, veracity-first, teamwork, and a focus on quality over quantity. During a normal work day, you will interact with critical infrastructure layers, from the Linux Kernel and JVM internals to cloud-wide orchestration.
YOU'LL BE RESPONSIBLE FOR
- Leading Deep-Dive Investigations: Conduct high-level performance analysis to identify and resolve systemic bottlenecks across our global JVM-based microservices architecture.
- Optimizing Resource Efficiency: Drive initiatives to reduce infrastructure costs and latency by fine-tuning JVM parameters, Garbage Collection (ZGC, G1), and memory management (heap and off-heap).
- Building Diagnostic Tooling: Develop and implement advanced observability tools using eBPF, JFR, and Flamegraphs to provide real-time insights into kernel and runtime behavior.
- Kernel & Runtime Alignment: Bridge the gap between the Linux Kernel and the JVM, optimizing thread scheduling (CFS/EEVDF) and managing resource isolation (cgroups/throttling) within our Kubernetes environment.
- Architecting Scalable Solutions: Design and deliver innovative infrastructure improvements that address long-term performance challenges, ensuring our systems scale ahead of demand.
- Technical Mentorship & Culture: Share expertise on JVM internals and performance best practices with the wider Engineering team, fostering a culture of technical excellence and "quality over quantity."
- Root Cause Excellence: Deep dive into complex concurrency issues, lock contentions, and memory leaks, providing definitive fixes for high-impact technical debt.
- Strategic Collaboration: Work closely with the Computing Squad to align orchestration strategies with system performance goals, ensuring a seamless interface between infrastructure and workloads.
WE ARE LOOKING FOR A PERSON WHO HAS
- Expertise in JVM Internals: Deep, low-level knowledge of the JVM is essential. You must understand how the JVM works "under the hood," including JIT compilation (C1/C2), class loading, and intrinsic methods.
- JVM Tuning & Garbage Collection: Extensive experience with GC algorithms (ZGC, G1, Shenandoah), including the ability to tune them for massive heaps and ultra-low latency requirements.
- OpenJDK Contribution (Major Plus): Previous experience contributing to the OpenJDK project or other low-level runtime environments is a significant advantage.
- Linux Kernel & Scheduling: Deep understanding of the Linux Scheduler (CFS/EEVDF), thread scheduling, and how the kernel manages high-concurrency Java workloads.
- Memory Architecture: Mastery of heap and off-heap memory management, including Direct Buffers, memory-mapped files, and diagnosing complex memory leaks.
- Advanced Diagnostics: Mastery of diagnostic tools such as Flamegraphs, JFR (Java Flight Recorder), eBPF, and performing large-scale heap dump analysis.
- Resource Isolation: Extensive experience with cgroups and the impact of CPU Throttling on JVM quotas within Kubernetes/EKS.
- Concurrency: Proven ability to diagnose and resolve complex concurrency problems, including lock contention and race conditions at the instruction level.
- Cloud Platforms: Knowledge of AWS infrastructure and its performance characteristics.
- Develops and delivers innovative solutions that address team-level or project-level challenges, focusing on medium and long-term impact
- Understand the technical aspects, capabilities, and limitations of our systems, contributing to discussions and improvements.
- Anticipate technical and product issues, making appropriate design decisions to avoid them
- Is enthusiastic about sharing knowledge and mentoring others.
- Deep dive into a problem to identify root causes when prioritized.
Benefits
- Opportunity of earning equity at Nu
- Medical Insurance
- Dental and Vision Insurance
- Life Insurance and AD&D
- Extended maternity and paternity leaves
- Nucleo - Our learning platform of courses
- NuLanguage - Our language learning program
- NuCare - Our mental health and wellness assistance program
- Extended maternity and paternity leaves
- 401K
- Saving Plans - Health Saving Account and Flexible Spending Account
- Work-from-home Allowance
- Relocation Assistance Package, if applicable.
Location for this opportunity (City, Country)
- Palo Alto, United States
- Miami, United States
- Washington DC, United States
- Durham, United States
WORK MODEL FOR THIS ROLE
Hybrid 2-3 times/week: Our hybrid work model brings us to the office at least twice a week, on strategic days designed to maximize team connection and collaboration. For more details, visit https://building.nubank.com/nu-hybrid-work-model/
Our recruitment process may involve the use of artificial intelligence–enabled tools, such as automated interview transcription and analysis, to support the evaluation process. Artificial intelligence is not used to make final hiring decisions; all decisions are made by human reviewers.