Video Annotator
Core
Watch short video clips of physical tasks, segment them into discrete steps, and write precise natural-language instructions describing the actions.
Role type
Video Annotator (Vision-Language-Action dataset)
Builds
VLA dataset with event-level annotations and step-by-step instructions
Domain
AI/ML data preparation, computer vision, robotics
Deliverable
production ML models
Required skills
English (C1+), attention to detail, written communication, video annotation tools, independent work, data consistency
Preferred skills
data annotation experience, AI/ML concepts (computer vision, NLP, robotics), instructional content writing, annotation platforms (CVAT, Label Studio, Scale)
Responsibilities
Segment videos into logical steps based on action boundaries, write clear and grammatically correct step descriptions, assign accurate timestamps, flag unsuitable videos, participate in calibration sessions, meet productivity and quality targets
Seniority
Individual Contributor