远程工作雷达

AI性能工程师

AI Performance Engineer

AI开发工程限定地区(需当地身份)
公司Bright Vision Technologies
薪资$75,000 - $100,000/年
工作地点United States
地域资格限定地区(需当地身份)
时区要求日间重叠约 9 小时,基本正常作息
用工类型Full Time
发布时间昨天
数据来源Himalayas
前往 Himalayas 查看并投递 →
注意地域限制:该职位明确限定在 United States 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

AI性能工程师 – 远程

Bright Vision Technologies 是一家技术咨询和软件开发公司,为美国各地提供云、AI、数据和企业解决方案。这是一个加入一家知名且备受尊敬的组织的绝佳机会,提供巨大的职业发展潜力。

职位名称:AI性能工程师
工作地点:100%远程(美国)
职位类型:全职,直接W2
薪资范围:每年75,000–100,000美元
所需经验:6年以上

赞助:美国公民、绿卡持有者、EAD持有者以及H-1B转签候选人欢迎申请。我们无法为该职位的新H-1B签证申请提供赞助。

职位简介
我们正在寻找一名AI性能工程师,专注于从大型神经网络系统中提取最大吞吐量、最小化延迟并降低成本。该角色涵盖从底层内核优化到分布式系统调优的整个技术栈,需要对GPU架构、模型并行性、内存管理以及编译器级优化有深入理解。理想的候选人应有在生产环境AI工作负载中产生实际影响的经验,具备强大的工具使用和测量能力,能够做出严谨的数据驱动优化决策。在此职位上,您将与跨职能团队——产品、设计、工程、运维和业务相关方密切合作,将模糊的需求转化为良好的工程解决方案,并期望通过代码审查、设计审查和对初级工程师的指导来提升标准。成功候选人应具备扎实的工程纪律、清晰的沟通风格以及在生产环境中交付高质量工作的记录。

必备资格
· 计算机科学、计算机工程或相关领域的学士或硕士学位。

  • 在性能工程、ML系统或HPC领域有6年或以上经验。
  • 精通Python和C++。
  • 有在现代GPU上优化深度学习工作负载的实际经验。
  • 深入了解分布式训练和推理技术。
  • 有在CPU、GPU和分布式系统上使用分析工具的经验。
  • 熟悉模型压缩技术及其准确性影响。
  • 对内存层次结构、通信原语和并行策略有深刻理解。
  • 出色的测量、调试和分析推理技能
查看英文原文

AI Performance Engineer – Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: AI Performance Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $75,000–$100,000 Annually
Experience Required: 6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary
We are seeking an AI Performance Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production.

Required Qualifications
· Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.

  • Six or more years of experience in performance engineering, ML systems, or HPC.
  • Strong proficiency in Python and C++.
  • Hands-on experience optimizing deep learning workloads on modern GPUs.
  • Deep understanding of distributed training and inference techniques.
  • Experience with profiling tools across CPU, GPU, and distributed systems.
  • Familiarity with model compression techniques and their accuracy implications.
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies.
  • Excellent measurement, debugging, and analytical reasoning skills.
  • Strong communication and collaboration skills.

Preferred Qualifications
· Experience optimizing LLM inference at production scale.

  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects.
  • Familiarity with custom kernel authoring in Triton or CUTLASS.
  • Experience with FinOps for AI workloads.
  • Publications or talks on AI systems performance.

How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to . Learn more about Bright Vision Technologies
Bright Vision Technologies is an Equal Opportunity Employer.Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

混合云架构师

Bright Vision TechnologiesUnited States$100,000 - $150,000/年Full Time昨天
开发工程限定地区(需当地身份)

云网络工程师

Bright Vision TechnologiesUnited States$100,000 - $150,000/年Full Time昨天
开发工程限定地区(需当地身份)

Sitecore 技术顾问

Bright Vision TechnologiesUnited States$100,000 - $150,000/年Full Time昨天
开发工程限定地区(需当地身份)

AI数据工程师

Bright Vision TechnologiesUnited States$80,000 - $100,000/年Full Time昨天
AI开发工程限定地区(需当地身份)

方案经理

Bright Vision TechnologiesUnited States$80,000 - $120,000/年Full Time昨天
市场运营职能支持限定地区(需当地身份)

← 返回全部职位