Project Cursa - Robot Manipulation Video Annotator (V2)

AI Trainer Jobs

United States

Remote

USD 40,000 - 60,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

AI Trainer Jobs is seeking detail-oriented annotators to label robot manipulation videos. You will describe segments with precise language, label actions at atomic, skill, and task levels, and ensure every moment is covered across three camera views.

You will follow a comprehensive style guide and participate in calibration sessions to align with client references. The role emphasizes strong written English, high attention to detail, and the ability to work independently on heads-down annotation

Qualifications

  • Strong written English with precise descriptive sentences.
  • Sharp attention to detail to distinguish small differences in grasps or moves.
  • Comfort following a detailed, structured style guide consistently.
  • Basic comfort with spatial/mechanical description (left/right, above/below, part names).
  • Reliable, self-directed work habits and ability to work heads-down.

Responsibilities

  • Watch robot manipulation videos from multiple camera angles and describe segments.
  • Write clear, natural-language descriptions for each segment.
  • Apply labels at three levels of detail for each applicable segment and ensure coverage of all moments.

Skills

Strong written English
Attention to detail
Structured style guide
Spatial description
Self-directed work

Tools

Label Studio

Job description

What You'll Do
  1. Watch short robot manipulation videos, each filmed from three synchronized camera views (an overhead view and views from each of the robot's two wrist-mounted cameras).
  2. Break each video into time segments and write clear, natural-language descriptions for each segment.
  3. Apply labels at three levels of detail for each applicable segment:
  • Atomic motion(a few seconds) — a single small movement (e.g., "close fingers around the red handle")
  • Skill / subtask(several seconds to ~20seconds) —a complete, meaningful action (e.g., "pick up the red block by its edge")
  • Task / goal(up to ~1minute) —the overall purpose of a sequence of skills (e.g., "place all blocks in the container")
  1. Ensure every moment of video is covered by a label at two or more of these levels — no gaps, including idle or pause moments.
  2. Accurately describe exactly what happens, including when something doesn't go as planned (a dropped object, a failed grasp, a slipped grip). Precision matters more than making the robot look successful.
  3. Cross-reference all three camera angles: use the overhead view to understand the overall scene and object identity, and the close-up wrist views to confirm exact contact and grasp details.
  4. Follow a detailed style guide covering vocabulary for actions, spatial relationships, object descriptions, and manner of movement, applying it consistently across many episodes.
  5. Participate in periodic calibration sessions to align your labeling with the team and the client's reference examples.

What We're Looking For
Required:

  1. Strong written English — you'll write dozens of short, precise descriptive sentences per video and need to vary your language rather than repeating the same phrases.
  2. Sharp attention to detail — able to distinguish small differences (a successful grasp vs. a fumble, a push vs. a drag, which specific object part is being touched).
  3. Comfort following a detailed, structured style guide and applying it consistently, even in ambiguous or edge-case scenarios.
  4. Basic comfort with spatial/mechanical description (left/right, above/below, naming object parts like handles, lids, or edges).
  5. Reliable, self-directed work habits — this is often heads-down work with periodic check-ins rather than close supervision.
Nice to Have
  1. Prior experience with video annotation, data labeling, transcription, or QA work.
  2. Familiarity with robotics terminology (grippers, end-effectors, manipulation) — helpful but not necessary, as the style guide is self-contained.
  3. Experience with annotation tools such as Label Studio.

#L1-CC1

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Project Cursa - Robot Manipulation Video Annotation QC
Project Cursa - Robot Manipulation Video Annotation QC

Annotation Academy • United States

Remote
USD 60,000 - 90,000
Project Cursa - Robot Manipulation Video Annotator (V2)
Project Cursa - Robot Manipulation Video Annotator (V2)

Annotation Academy • United States

Remote
USD 21,000 - 30,000
Robotics Video Annotator: Precise Manipulation Labels
Robotics Video Annotator: Precise Manipulation Labels

AI Trainer Jobs • United States

Remote
USD 40,000 - 60,000
Robotics Expert – Remote
Robotics Expert – Remote

RemoteJobsOne • Charlotte (NC)

Remote
USD 69,000 - 124,000
Remote Robotic Arm Video Annotator & QA Specialist
Remote Robotic Arm Video Annotator & QA Specialist

OpenTrain AI, Inc. • Northern (KY)

Hybrid
USD 69,000 - 124,000
Robotic Arm Video Annotation Editor
Robotic Arm Video Annotation Editor

OpenTrain AI, Inc. • United States

Remote
USD 69,000 - 124,000
Remote Robot Manipulation Video Annotator
Remote Robot Manipulation Video Annotator

Annotation Academy • United States

Remote
USD 21,000 - 30,000
Video Content Moderator - Remote
Video Content Moderator - Remote

YO AI Labs • North Carolina

Remote
USD 28,000 - 50,000
Video Content Reviewer - Remote
Video Content Reviewer - Remote

YO AI Labs • North Carolina

Remote
USD 28,000 - 55,000
Video Annotation Specialist - Remote
Video Annotation Specialist - Remote

YO AI Labs • North Carolina

Remote
USD 34,000 - 69,000