×
Register Here to Apply for Jobs or Post Jobs. X

Project Cursa - Robot Manipulation Video Annotator; V2

Job in Houston, Harris County, Texas, 77246, USA
Listing for: AI Trainer Jobs
Full Time position
Listed on 2026-10-09
Job specializations:
  • Creative Arts/Media
Salary/Wage Range or Industry Benchmark: 40000 - 60000 USD Yearly USD 40000.00 60000.00 YEAR
Job Description & How to Apply Below
Position: Project Cursa - Robot Manipulation Video Annotator (V2)
What You'll Do
  • Watch short robot manipulation videos, each filmed from three synchronized camera views (an overhead view and views from each of the robot's two wrist-mounted cameras).
  • Break each video into time segments and write clear, natural-language descriptions for each segment.
  • Apply labels at three levels of detail for each applicable segment:
    • Atomic motion(a few seconds) — a single small movement (e.g., "close fingers around the red handle")
    • Skill / subtask(several seconds to ~20seconds) —a complete, meaningful action (e.g., "pick up the red block by its edge")
    • Task / goal(up to ~1minute) —the overall purpose of a sequence of skills (e.g., "place all blocks in the container")
  • Ensure every moment of video is covered by a label at two or more of these levels — no gaps, including idle or pause moments.
  • Accurately describe exactly what happens, including when something doesn't go as planned (a dropped object, a failed grasp, a slipped grip). Precision matters more than making the robot look successful.
  • Cross-reference all three camera angles: use the overhead view to understand the overall scene and object identity, and the close-up wrist views to confirm exact contact and grasp details.
  • Follow a detailed style guide covering vocabulary for actions, spatial relationships, object descriptions, and manner of movement, applying it consistently across many episodes.
  • Participate in periodic calibration sessions to align your labeling with the team and the client's reference examples.
  • What We're Looking For

    Required:

  • Strong written English — you'll write dozens of short, precise descriptive sentences per video and need to vary your language rather than repeating the same phrases.
  • Sharp attention to detail — able to distinguish small differences (a successful grasp vs. a fumble, a push vs. a drag, which specific object part is being touched).
  • Comfort following a detailed, structured style guide and applying it consistently, even in ambiguous or edge-case scenarios.
  • Basic comfort with spatial/mechanical description (left/right, above/below, naming object parts like handles, lids, or edges).
  • Reliable, self-directed work habits — this is often heads-down work with periodic check-ins rather than close supervision.
  • Nice to Have
  • Prior experience with video annotation, data labeling, transcription, or QA work.
  • Familiarity with robotics terminology (grippers, end-effectors, manipulation) — helpful but not necessary, as the style guide is self-contained.
  • Experience with annotation tools such as Label Studio.
  • #L1-CC1

    To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
    (If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
     
     
     
    Search for further Jobs Here:
    (Try combinations for better Results! Or enter less keywords for broader Results)
    Location
    Increase/decrease your Search Radius (miles)
    0
    200
    Filters
    Education Level
    Experience Level (years)
    Posted in last:
    Salary