远程工作雷达

AI安全专员 - 评估专家

AI Safety Specialist - Evaluation Expert

AI市场运营未标注地域日间重叠约 2 小时,需偶尔早起或晚睡
公司mercor
薪资$60 - $70
工作地点Kosovo
地域资格未标注地域
时区要求日间重叠约 2 小时,需偶尔早起或晚睡
用工类型Contractor
发布时间昨天
数据来源Himalayas
前往 Himalayas 查看并投递 →
作息提示:日间重叠约 2 小时,需偶尔早起或晚睡。

关于职位
Mercor将顶尖的创意和技术人才与领先的AI研究实验室联系起来。公司总部位于旧金山,我们的投资人包括Benchmark、General Catalyst、Peter Thiel、Adam D'Angelo、Larry Summers和Jack Dorsey。
职位:AI安全实践者
类型:合同制
薪酬:60-70美元/小时
地点:远程办公
岗位职责

  • 评估AI生成的回答在安全、事实准确性、政策合规性和整体质量方面的表现。
  • 审查涉及虚假信息、政治宣传、自残、暴力、网络、生物安全和其他敏感领域的相关内容。
  • 应用并优化RLHF、SFT和AI安全基准测试的评估标准。
  • 识别不安全的输出、幻觉、推理失败和政策违规。
  • 提供结构化反馈以提升模型对齐度和安全性能。
  • 与AI研究人员和安全团队合作,开展持续的评估项目。

任职要求

必备条件

  • 新闻学、传播学、心理学、社会学、公共政策、法律、生物学、化学、计算机科学或相关专业的学士及以上学位。
  • 在AI安全、信任与安全、新闻业、公共政策、科研、安全或相关领域有5年以上专业经验。
  • 优秀的英文书面表达能力、批判性思维和分析推理能力。
  • 能够持续评估复杂且政策敏感的场景。

优先考虑

  • 具有AI安全、RLHF、SFT、信任与安全或AI评估方面的经验。
  • 熟悉安全政策、内容审核或评估标准开发。
  • 具有审查复杂、高风险或模糊内容的经验。

申请流程(需要20-30分钟完成)

  • 上传简历
  • 基于简历的AI面试
  • 提交表单

资源与支持

  • 有关面试流程和平台信息的详细信息,请查看:
  • 如有任何帮助或支持,请联系:

备注:我们的团队每天都会审核申请。请完成AI面试和申请步骤,以便考虑此机会。
最初发布于Himalayas

查看英文原文

About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.
Position: AI Safety Practitioner
Type:Contract
Compensation:$60–$70/hour
Location:Remote
Role Responsibilities

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
  • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
  • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
  • Provide structured feedback to improve model alignment and safety performance.
  • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Qualifications

Must-Have

  • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.
  • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field.
  • Excellent written English, critical thinking, and analytical reasoning skills.
  • Ability to consistently evaluate nuanced and policy-sensitive scenarios.

Preferred

  • Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
  • Familiarity with safety policies, content moderation, or evaluation rubric development.
  • Experience reviewing complex, high-risk, or ambiguous content.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

  • For details about the interview process and platform information, please check:
  • For any help or support, reach out to:

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
Originally posted on Himalayas

本页面信息整理自 Himalayas,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位