资深机器学习工程师, AI 安全
Staff Machine Learning Engineer, AI Security
Reddit 是一个由社区组成的社区。它建立在共同的兴趣、热情和信任之上,是互联网上最开放和真实的对话场所。每天,Reddit 用户提交、投票并评论他们最关心的话题。拥有 100,000 多个活跃社区,以及约 1.3 亿日活跃独立访客,Reddit 是互联网上最大的信息来源之一。如需更多信息,请访问 www.redditinc.com。
Reddit 安全平台工程团队的 AI 安全小组构建安全机制,将安全融入 Reddit 的产品、工程系统、AI 平台和运营基础设施,使安全路径成为对人和代理最便捷的路径。这项工作的一个核心部分是开发实用且高质量的机器学习系统,以检测和防止诸如提示注入、越狱、敏感数据泄露以及不安全或未经授权的 AI 行为等风险。基于 Reddit 的集中式 LLM 安全护栏平台,该平台在主要产品中提供共享的安全与安全边界,我们正在扩展 ML 驱动的保护措施,以适应 Reddit 的 AI 系统和威胁环境的变化。
我们正在寻找一位创始级资深机器学习工程师,负责 Reddit AI 安全领域的模型开发、训练和优化。这是一个战略性和亲力亲为的个人贡献者角色,结合对模型架构、训练数据和实验的深度掌控,以及跨团队的技术领导力。你将帮助 Reddit 提供更强大的 AI 保护措施,同时保持高质量的产品体验。
你将产生的影响
- 选择、适配、微调、评估和部署针对 Reddit 特定安全问题的预训练模型和轻量级分类器。
- 在 Reddit 的 ML 平台上构建可复现的训练和评估流程,与平台工程师合作提升推理性能、资源效率和操作可靠性。
- 制定技术愿景和多季度建模路线图,与跨职能团队合作收集需求、定义模型架构,并迭代模型开发。
- 进行模型评估和性能分析,以提高准确性和对抗鲁棒性,并定义平衡误报率、延迟、吞吐量、可靠性和成本的上线标准。
- 负责训练数据质量和生产模型生命周期,利用监控、事件发现和红队反馈来指导数据集改进。
查看英文原文
Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 130 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com.
The AI Security team within Reddit’s Security Platform Engineering organization builds security into Reddit’s products, engineering systems, AI platforms, and operational infrastructure so the secure path is the easiest path for both people and agents. A core part of this work is developing practical, high-quality machine learning systems that detect and prevent risks such as prompt injection, jailbreaks, sensitive data exposure, and unsafe or unauthorized AI behavior. Building on Reddit’s centralized LLM Guardrails Platform, which provides a shared security and safety boundary across major products, we are expanding ML-powered protections as Reddit’s AI systems and threat landscape evolve.
We are looking for a founding Staff Machine Learning Engineer to lead the development, training, and optimization of models for AI security at Reddit. This is a strategic and hands-on individual contributor role, combining deep ownership of model architecture, training data, and experimentation with technical leadership across teams. You will help Reddit deliver stronger AI protections while preserving a high-quality product experience.
How You’ll Have Impact
- Select, adapt, fine-tune, evaluate, and deploy pretrained models and lightweight classifiers for Reddit-specific security problems.
- Build reproducible training and evaluation pipelines on Reddit’s ML platform, partnering with platform engineers to improve inference performance, resource efficiency, and operational reliability.
- Set the technical vision and multi-quarter modeling roadmap, partnering with cross-functional teams to gather requirements, define model architectures, and iterate on model development.
- Conduct model evaluations and performance analysis to improve accuracy and adversarial robustness, and define launch criteria that balance false positives, latency, throughput, reliability, and cost.
- Own training-data quality and the production model lifecycle, using monitoring, incident findings, and red-team feedback to guide dataset improvements, retraining, and safe rollout or rollback.
- Establish best practices for responsible ML development and deployment, including reproducible experiments, testing, model and data lineage, and privacy-aware data use.
- Stay current with research in NLP, large language models, and relevant multimodal techniques, translating promising advances into measurable model improvements.
- Mentor engineers and lead technical discussions and reviews, shaping the team’s long-term ML capabilities and AI security direction.
Who You Might Be
- 8+ years of experience developing machine learning models, with substantial hands-on model training experience, demonstrated production impact, and a record of leading complex initiatives across teams.
- Strong background in Python programming, software engineering, and deep learning frameworks and libraries such as TensorFlow, PyTorch, or Hugging Face Transformers.
- Deep understanding of neural network architectures and optimization, with proficiency in data preprocessing, tokenization, embeddings, language modeling, and model calibration.
- Expertise in scalable data pipelines and distributed training frameworks such as Ray Train or PyTorch Distributed, with a strong understanding of hardware and system tradeoffs.
- Demonstrated rigor in experimental design and model evaluation, including representative holdouts, ablation studies, adversarial tests, and error analysis to diagnose training issues, bias, and generalization gaps.
- Excellent written and verbal communication, with the ability to explain model behavior, security risk, uncertainty, and tradeoffs to technical and non-technical partners.
Preferred Qualifications
- Experience applying ML to security, trust and safety, fraud, privacy, or related adversarial domains.
- Experience with adversarial training, model distillation, active learning, or synthetic-data generation to improve model quality, robustness, and training efficiency.
Benefits:
- Comprehensive Healthcare Benefits and Income Replacement Programs
- 401k with Employer Match
- Global Benefit programs that fit your lifestyle, from workspace to professional development to caregiving support
- Family Planning Support
- Gender-Affirming Care
- Mental Health & Coaching Benefits
- Flexible Vacation & Paid Volunteer Time Off
- Generous Paid Parental Leave
#LI-Remote
Pay Transparency:
This job posting may span more than one career level.
In addition to base salary, this job is eligible to receive equity in the form of restricted stock units, and depending on the position offered, it may also be eligible to receive a commission. Additionally, Reddit offers a wide range of benefits to U.S.-based employees, including medical, dental, and vision insurance, 401(k) program with employer match, generous time off for vacation, and parental leave. To learn more, please visit https://www.redditinc.com/careers/.
To provide greater transparency to candidates, we share base salary ranges for all US-based job postings regardless of state. We set standard base pay ranges for all roles based on function, level, and country location, benchmarked against similar stage growth companies. Final offer amounts are determined by multiple factors including, skills, depth of work experience and relevant licenses/credentials, and may vary from the amounts listed below.
The base salary range for this position is:
$230,000—$322,000 USD
In select roles and locations, the interviews will be recorded, transcribed and summarized by artificial intelligence (AI). You will have the opportunity to opt out of recording, transcription and summarization prior to any scheduled interviews.
During the interview, we will collect the following categories of personal information: Identifiers, Professional and Employment-Related Information, Sensory Information (audio/video recording), and any other categories of personal information you choose to share with us. We will use this information to evaluate your application for employment or an independent contractor role, as applicable. We will not sell your personal information or disclose it to any third party for their marketing purposes. We will delete any recording of your interview promptly after making a hiring decision. For more information about how we will handle your personal information, including our retention of it, please refer to our Candidate Privacy Policy for Potential Employees and Contractors.
Reddit is proud to be an equal opportunity employer, and is committed to building a workforce representative of the diverse communities we serve. Reddit is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If, due to a disability, you need an accommodation during the interview process, please let your recruiter know.