AI安全专家 - 红队
AI Safety Expert - Red Teaming
关于职位
Mercor 将顶尖的创意和技术人才与领先的 AI 研究实验室联系起来。公司总部位于旧金山,我们的投资人包括 Benchmark、General Catalyst、Peter Thiel、Adam D'Angelo、Larry Summers 和 Jack Dorsey。
职位:AI 安全专家 — 英语 & 奥里亚语
类型:合同制
薪酬:20–22 美元/小时
地点:远程
岗位职责
- 对对话式 AI 模型和代理进行红队测试,以识别越狱、提示注入和滥用案例。
- 通过标注失败案例、分类漏洞和标记系统性风险,生成高质量的人类数据。
- 通过遵循分类法、基准测试和操作手册来建立结构,以保持测试的一致性。
- 可重复记录,以生成报告、数据集和攻击案例供客户采取行动。
- 独立且异步工作,以满足截止日期并提升 AI 模型性能。
任职要求
必备条件
- 必须具备流利的语言能力:英语 & 奥里亚语。
- 在 AI 对抗工作、网络安全或社会技术探测方面有红队经验。
- 具备良好的沟通能力,能够向技术人员和非技术人员解释风险。
优先考虑
- 在对抗机器学习、网络安全或社会技术风险方面的经验。
- 具备创造性探测技能,如心理学、表演或写作,用于非常规的对抗性思维。
申请流程(需要 20–30 分钟完成)
- 上传简历
- 基于简历的 AI 面试
- 提交表单
资源与支持
- 有关面试流程和平台信息的详细信息,请查看:
- 如有任何帮助或支持,请联系:
备注:我们的团队每天都会审核申请。请完成 AI 面试和申请步骤,以便考虑此机会。
最初发布于 喜马拉雅山
查看英文原文
About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.
Position: AI Safety Experts — English & Odia
Type:Contract
Compensation:$20–$22/hour
Location:Remote
Role Responsibilities
- Red team conversational AI models and agents to identify jailbreaks, prompt injections, and misuse cases.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing.
- Document reproducibly to produce reports, datasets, and attack cases for customer action.
- Work independently and asynchronously to meet deadlines while improving AI model performance.
Qualifications
Must-Have
- Fluent Language Skills Required: English & Odia.
- Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
- Strong communication skills to explain risks to technical and non-technical stakeholders.
Preferred
- Experience in Adversarial ML, Cybersecurity, or Socio-technical risk.
- Skills in Creative probing such as psychology, acting, or writing for unconventional adversarial thinking.
Application Process (Takes 20–30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form
Resources & Support
- For details about the interview process and platform information, please check:
- For any help or support, reach out to:
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
Originally posted on Himalayas