远程工作雷达

项目Cursa - 机器人操作视频标注员

Project Cursa - Robot Manipulation Video Annotator

AI限定地区(需当地身份)
公司Weloglobal
薪资未公开
工作地点Philippines
地域资格限定地区(需当地身份)
时区要求无特别要求
用工类型Freelance
发布时间未知
数据来源Lever
前往企业招聘页投递 →
注意地域限制:该职位明确限定在 Philippines 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

职位描述

我们正在寻找注重细节的标注员,帮助为AI训练目的对机器人操作视频进行标注。你将观看机器人执行操作任务的短片(从三个同步摄像机角度拍摄),并为所发生的行为生成精确、结构化、自然语言的描述。这份工作直接支持机器人AI模型的开发,需要较强的英文书面表达能力、敏锐的观察力,以及持续遵循详细风格指南的纪律性。

你将要做的事情

  • 观看短时机器人操作视频,每个视频从三个同步摄像机视角拍摄(一个俯视视角和机器人两个腕部摄像头的视角)。
  • 将每个视频分成时间片段,并为每个片段撰写清晰的自然语言描述。
  • 为每个适用的片段应用三个层级的标签:

·

  • 原子动作(几秒钟)—— 一个小型动作(例如:“手指围绕红色把手闭合”)
  • 技能/子任务(几秒到约20秒)—— 一个完整且有意义的动作(例如:“通过边缘拿起红色方块”)
  • 任务/目标(最多约1分钟)—— 一系列技能的总体目标(例如:“将所有方块放入容器中”)
  • 确保视频中的每一刻都被至少两个层级的标签覆盖——没有遗漏,包括空闲或暂停时刻。
  • 准确描述实际发生的情况,包括当事情未按计划进行时(如物品掉落、抓取失败、握持滑脱)。精确性比让机器人看起来成功更重要。
  • 交叉参考所有三个摄像机视角:使用俯视图来理解整体场景和物体身份,使用近距离腕部视角来确认具体的接触和抓取细节。
  • 遵循详细的风格指南,涵盖动作词汇、空间关系、物体描述和运动方式,且在多个剧集上一致应用。
  • 参与定期校准会议,使你的标注与团队和客户的参考示例保持一致。

我们寻找的人选

要求:

  • 出色的英文书面表达能力——你将为每个视频撰写数十个简短、精确的描述性句子,需要使用多样的语言而不是重复相同的短语。
  • 敏锐的细节关注能力——能够区分细微差别(成功的抓取与失误、推动与拖动、具体哪个物体部分被触碰)。
  • 舒适于长时间专注工作,能处理重复性任务并保持高精度。
查看英文原文

About the Role

We are looking for detail-oriented annotators to help label robot manipulation videos for AI training purposes. You'll watch short videos of robots performing manipulation tasks (filmed from three synchronized camera angles) and produce precise, structured, natural-language descriptions of the actions taking place. This work directly supports the development of robotics AI models and requires strong written English, sharp observational skills, and the discipline to follow a detailed style guide consistently.

What You'll Do

  • Watch short robot manipulation videos, each filmed from three synchronized camera views (an overhead view and views from each of the robot's two wrist-mounted cameras).
  • Break each video into time segments and write clear, natural-language descriptions for each segment.
  • Apply labels at three levels of detail for each applicable segment:

·

  • Atomic motion (a few seconds) — a single small movement (e.g., "close fingers around the red handle")
  • Skill / subtask (several seconds to ~20 seconds) — a complete, meaningful action (e.g., "pick up the red block by its edge")
  • Task / goal (up to ~1 minute) — the overall purpose of a sequence of skills (e.g., "place all blocks in the container")
  • Ensure every moment of video is covered by a label at two or more of these levels — no gaps, including idle or pause moments.
  • Accurately describe exactly what happens, including when something doesn't go as planned (a dropped object, a failed grasp, a slipped grip). Precision matters more than making the robot look successful.
  • Cross-reference all three camera angles: use the overhead view to understand the overall scene and object identity, and the close-up wrist views to confirm exact contact and grasp details.
  • Follow a detailed style guide covering vocabulary for actions, spatial relationships, object descriptions, and manner of movement, applying it consistently across many episodes.
  • Participate in periodic calibration sessions to align your labeling with the team and the client's reference examples.

What We're Looking For

Required:

  • Strong written English — you'll write dozens of short, precise descriptive sentences per video and need to vary your language rather than repeating the same phrases.
  • Sharp attention to detail — able to distinguish small differences (a successful grasp vs. a fumble, a push vs. a drag, which specific object part is being touched).
  • Comfort following a detailed, structured style guide and applying it consistently, even in ambiguous or edge-case scenarios.
  • Basic comfort with spatial/mechanical description (left/right, above/below, naming object parts like handles, lids, or edges).
  • Reliable, self-directed work habits — this is often heads-down work with periodic check-ins rather than close supervision.

Nice to Have:

  • Prior experience with video annotation, data labeling, transcription, or QA work.
  • Familiarity with robotics terminology (grippers, end-effectors, manipulation) — helpful but not necessary, as the style guide is self-contained.
  • Experience with annotation tools such as Label Studio.

Why Join Welo Data?

✨ Limitless Flexibility

Project-based opportunities that fit your availability. Choose when and how much you want to contribute—fully remote, with complete autonomy.

🌱 Limitless Growth

Optional access to AI and Large Language Model workshops designed specifically for professionals like you. No coding required—just your expertise.

🌍 Limitless Support

Be part of a global contributor community with responsive guidance and support.

💡 Real Impact

Apply your expertise in the Legal field to influence the AI systems shaping the future of your industry—while collaborating with data professionals and expanding your skills.

How to Apply?

Apply now by answering a few quick questions to join our database and become part of our growing community.

About Welo Data

Welo Data, part of Welocalize, is a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train the world’s most advanced AI systems. We’re building smarter, more human AI with a diverse community in 100+ countries.

At Welo Data, Limitless AI. Limitless You. isn’t just a slogan—it’s our promise. We build smarter AI through the power of human contribution, offering limitless opportunities for our global community to grow, contribute, and work on their terms.

本页面信息整理自 Lever,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位