AI研究工程师(模型压缩与量化)
AI Research Engineer (Model Compression & Quantization)
Tags: Web3 职位 • Web3 研究职位 • Web3 AI 职位 • Web3 机器学习职位 • 区块链远程职位 • Web3 全职职位 • Web3 职位 • 区块链稳定币职位
加入 Tether,塑造数字金融的未来。在 Tether,我们不仅仅是构建产品,更是在引领一场全球金融革命。我们前沿的解决方案使企业——从交易所、钱包到支付处理器和自动取款机——能够无缝地在区块链上集成储备支持的代币。通过利用区块链技术的力量,Tether 让您能够以极低的成本,即时、安全且全球性地存储、发送和接收数字代币。透明度是我们一切工作的基石,确保每笔交易都值得信赖。与 Tether 一起创新。
Tether Finance:我们的创新产品组合包括全球最值得信赖的稳定币 USDT,全球数亿人依赖它,同时还提供开创性的数字资产代币化服务。但这只是开始:
Tether Power:推动可持续增长,我们的能源解决方案使用环保方法,优化比特币挖矿的多余电力,位于现代化、地理分布广泛的设施中。
Tether Data:推动人工智能和点对点技术的突破,我们通过 KEET 等前沿解决方案降低基础设施成本,提升全球通信效率,这款旗舰应用重新定义了安全和私密的数据共享。
Tether Education:让顶级数字学习触手可及,我们赋能个人在数字和零工经济中蓬勃发展,推动全球增长和机遇。
Tether Evolution:在科技与人类潜能的交汇点,我们不断突破可能的边界,打造一个创新与人类能力融合的未来,这种融合前所未有。
为什么加入我们?我们的团队是全球人才的强大力量,成员来自世界各地的远程办公。如果你热衷于在金融科技领域留下印记,这就是你与最聪明的头脑合作、突破界限并制定新标准的机会。我们发展迅速,保持精简,并在行业中确立了领导地位。如果你具备出色的英语沟通能力,并准备好为地球上最具创新性的平台做出贡献,Tether 就是你的选择。你准备好成为未来的一部分了吗?
关于职位:作为我们 AI 研究团队的一员,你将推动模型压缩和高效部署方面的创新。
查看英文原文
Tags: Web3 Jobs • Cryptocurrency Research Jobs • Cryptocurrency Ai Jobs • Cryptocurrency Machine Learning Jobs • Web3 Remote Jobs • Cryptocurrency Full Time Jobs • Cryptocurrency Web3 Jobs • Web3 Stablecoin JobsJoin Tether and Shape the Future of Digital Finance. At Tether, we’re not just building products, we’re pioneering a global financial revolution. Our cutting-edge solutions empower businesses—from exchanges and wallets to payment processors and ATMs—to seamlessly integrate reserve-backed tokens across blockchains. By harnessing the power of blockchain technology, Tether enables you to store, send, and receive digital tokens instantly, securely, and globally, all at a fraction of the cost. Transparency is the bedrock of everything we do, ensuring trust in every transaction. Innovate with Tether. Tether Finance: Our innovative product suite features the world’s most trusted stablecoin, USDT, relied upon by hundreds of millions worldwide, alongside pioneering digital asset tokenization services. But that’s just the beginning: Tether Power: Driving sustainable growth, our energy solutions optimize excess power for Bitcoin mining using eco-friendly practices in state-of-the-art, geo-diverse facilities. Tether Data: Fueling breakthroughs in AI and peer-to-peer technology, we reduce infrastructure costs and enhance global communications with cutting-edge solutions like KEET, our flagship app that redefines secure and private data sharing. Tether Education: Democratizing access to top-tier digital learning, we empower individuals to thrive in the digital and gig economies, driving global growth and opportunity. Tether Evolution: At the intersection of technology and human potential, we are pushing the boundaries of what is possible, crafting a future where innovation and human capabilities merge in powerful, unprecedented ways. Why Join Us? Our team is a global talent powerhouse, working remotely from every corner of the world. If you’re passionate about making a mark in the fintech space, this is your opportunity to collaborate with some of the brightest minds, pushing boundaries and setting new standards. We’ve grown fast, stayed lean, and secured our place as a leader in the industry. If you have excellent English communication skills and are ready to contribute to the most innovative platform on the planet, Tether is the place for you. Are you ready to be part of the future? About the job. As a member of our AI research team, you will drive innovation in model compression and efficient deployment for advanced multimodal AI systems, including large language models (LLMs) and vision-language models (VLMs). Your work will focus on reducing model footprint and computational cost while preserving accuracy, enabling high-performance AI to run efficiently across resource-constrained edge devices. You will apply and advance compression techniques such as quantization, knowledge distillation, and pruning to streamline complex multimodal architectures that integrate text, images, and audio. We expect you to have deep expertise in model compression methods and a strong background in multimodal model architectures. You will adopt a hands-on, research-driven approach to develop, test, and implement novel compression strategies that balance model size, latency, throughput, and accuracy. Your responsibilities include building robust compression pipelines, establishing performance and fidelity metrics, and addressing bottlenecks in production inference. The ultimate goal is to deliver scalable, low-memory, low-latency AI systems on edge devices (i.e., smartphones) that maintain high fidelity and tangible real-world value. Responsibilities. Apply low-bit quantization to reduce model size and inference latency for generative AI models (LLMs, VLMs, multimodal) while maintaining accuracy and output quality. Leverage knowledge distillation to transfer capabilities from larger teacher models to smaller student models, enabling efficient multimodal reasoning across text, image, and audio inputs. Implement pruning techniques to remove redundant parameters and attention heads, reducing computational overhead without sacrificing task performance. Analyze trade-offs between model efficiency (size, latency, memory) and accuracy across quantization, distillation, and pruning methods; propose improvements based on empirical findings. Research and apply mixed-precision quantization and other advanced compression strategies (e.g., adaptive pruning schedules, distillation with intermediate feature matching) to optimize the accuracy–performance balance. Stay current with the latest research in model compression, including emerging techniques for multimodal and generative architectures. Document methodologies, experiments, and results clearly to support reproducibility, internal collaboration, and stakeholder communication. Author technical papers and publish findings in top-tier conferences (e.g., NeurIPS, ICML, ICLR, CVPR, ACL, AAAI) to advance the field of model compression for multimodal AI.Apply here 👉 AI Research Engineer (Model Compression & Quantization)