远程工作雷达

AI推理工程师 QVAC(全球远程)

AI Inference Engineer QVAC (100% remote Worldwide)

AI开发工程全球可投(据职位描述推断)
公司Tether Operations Limited
薪资未公开
工作地点Worldwide
地域资格全球可投(据职位描述推断)
时区要求日间重叠约 9 小时,基本正常作息
用工类型Full Time
发布时间今天
数据来源Himalayas
前往 Himalayas 查看并投递 →
全球可投:该职位未限制候选人所在地区。仍需注意薪资可能按地区折算,以及实际签约方式(正式雇佣 / 独立合同)。

加入Tether,塑造数字金融的未来
在Tether,我们不仅仅是构建产品,更是在引领一场全球金融革命。我们前沿的解决方案赋能企业——从交易所、钱包到支付处理器和ATM机——无缝地在区块链上集成储备支持的代币。通过利用区块链技术的力量,Tether使您能够以极低的成本,即时、安全、全球地存储、发送和接收数字代币。透明度是我们一切工作的基石,确保每笔交易都值得信赖。
与Tether一起创新
Tether Finance:我们的创新产品套件包括全球最值得信赖的稳定币USDT,全球数亿人依赖它,同时还提供开创性的数字资产代币化服务。
但这只是开始:
Tether Power:推动可持续增长,我们的能源解决方案利用环保实践,优化多余电力用于比特币挖矿,采用最先进的、地理分布广泛的设施。
Tether Data:推动人工智能和点对点技术的突破,我们通过KEET等前沿解决方案降低基础设施成本,提升全球通信效率,重新定义安全和私密的数据共享。
Tether Education:普及顶级数字学习资源,赋能个人在数字和零工经济中蓬勃发展,推动全球增长和机遇。
Tether Evolution:在技术和人类潜能的交汇点,我们不断突破可能的边界,打造一个创新与人类能力融合的未来。
为什么加入我们?
我们的团队是全球人才的强大力量,来自世界各地的远程办公。如果您热衷于在金融科技领域留下印记,这是您与最聪明的头脑合作、突破界限并制定新标准的机会。我们发展迅速,保持精简,并在行业中确立了领导地位。
如果您具备出色的英语沟通能力,并准备好为地球上最具创新性的平台做出贡献,Tether就是您的选择。
您准备好成为未来的一部分了吗?
职位简介:
您将负责QVAC本地AI堆栈背后的推理基础架构:使模型在真实用户硬件上快速、可靠且可预测运行的C++系统层。该职位的核心是运行时级别的工程品质,包括启动行为、内存压力、吞吐量等。

查看英文原文

Join Tether and Shape the Future of Digital Finance
At Tether, we’re not just building products, we’re pioneering a global financial revolution. Our cutting-edge solutions empower businesses—from exchanges and wallets to payment processors and ATMs—to seamlessly integrate reserve-backed tokens across blockchains. By harnessing the power of blockchain technology, Tether enables you to store, send, and receive digital tokens instantly, securely, and globally, all at a fraction of the cost. Transparency is the bedrock of everything we do, ensuring trust in every transaction.
Innovate with Tether
Tether Finance: Our innovative product suite features the world’s most trusted stablecoin, USDT, relied upon by hundreds of millions worldwide, alongside pioneering digital asset tokenization services.
But that’s just the beginning:
Tether Power: Driving sustainable growth, our energy solutions optimize excess power for Bitcoin mining using eco-friendly practices in state-of-the-art, geo-diverse facilities.
Tether Data: Fueling breakthroughs in AI and peer-to-peer technology, we reduce infrastructure costs and enhance global communications with cutting-edge solutions like KEET, our flagship app that redefines secure and private data sharing.
Tether Education: Democratizing access to top-tier digital learning, we empower individuals to thrive in the digital and gig economies, driving global growth and opportunity.
Tether Evolution: At the intersection of technology and human potential, we are pushing the boundaries of what is possible, crafting a future where innovation and human capabilities merge in powerful, unprecedented ways.
Why Join Us?
Our team is a global talent powerhouse, working remotely from every corner of the world. If you’re passionate about making a mark in the fintech space, this is your opportunity to collaborate with some of the brightest minds, pushing boundaries and setting new standards. We’ve grown fast, stayed lean, and secured our place as a leader in the industry.
If you have excellent English communication skills and are ready to contribute to the most innovative platform on the planet, Tether is the place for you.
Are you ready to be part of the future?
About the role:
You will own the inference backbone behind QVAC's local AI stack: the C++ systems layer that makes models run fast, reliably, and predictably on real user hardware. The role is centered on engineering quality at runtime level, including startup behavior, memory pressure, throughput/latency balance, and long-session stability. You will define and evolve the core abstractions that inference features depend on, so new capabilities can be added without sacrificing performance or maintainability. This is a role for someone who enjoys low-level problem solving, clear technical ownership, and building infrastructure that other teams trust in production. Your work directly enables private, on-device AI experiences and helps set the technical foundation for QVAC's next generation of peer-to-peer AI products.
About the job
You'll work on the C++ layer that powers local AI, porting and enhancing inference engines like llama.cpp or similar, to run efficiently on edge devices. Your focus is on the runtime: making models load faster, run leaner, and perform well across different hardware. You'll ensure that the inference layer is stable, optimized, and ready for integration with the rest of the stack.
This role is for engineers who want to work close to the metal, enabling private and fast on-device AI without relying on cloud infrastructure.
Responsibilities

  • Work on deploying machine learning models to edge devices using the frameworks: llama.cpp, ggml
  • Collaborate closely with researchers to assist in coding, training and transitioning models from research to production environments
  • Integrate AI features into existing products, enriching them with the latest advancements in machine learning

Requirements

  • Strong programming skills in C++
  • Strong experience with Llama.cpp and ggml inference engines, which facilitates the deployment of models to specific GPU architectures
  • Experience with any GPU framework between Cuda, Vulkan, Metal, OpenCL
  • Good understanding of deep learning concepts and model architectures
  • Experience with transformers, LLMs, Diffusion models
  • Demonstrated ability to rapidly assimilate new technologies and techniques
  • A degree in Computer Science, AI, Machine Learning, or a related field, complemented by a solid track record in AI R&D

Bonus points if:

  • You know how to train/fine-tune a LLM
  • You have productionized models
  • You have research experience in new model architectures
  • You have experience with distributed systems
  • Javascript experience

Important information for candidates
Recruitment scams have become increasingly common. To protect yourself, please keep the following in mind when applying for roles:

  • Apply only through our official channels. We do not use third-party platforms or agencies for recruitment unless clearly stated. All open roles are listed on our official careers page:
  • Verify the recruiter’s identity. All our recruiters have verified LinkedIn profiles. If you’re unsure, you can confirm their identity by checking their profile or contacting us through our website.
  • Be cautious of unusual communication methods. We do not conduct interviews over WhatsApp, Telegram, or SMS. All communication is done through official company emails and platforms.
  • Double-check email addresses. All communication from us will come from emails ending in @ or @
  • We will never request payment or financial details. If someone asks for personal financial information or payment at any point during the hiring process, it is a scam. Please report it immediately.

When in doubt, feel free to reach out through our official website.
Highlights
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位