About this opportunity
About the Role
We are looking for detail-oriented Video Annotators to support a Vision-Language-Action (VLA) dataset project. You will review short videos of physical tasks and activities, divide them into meaningful steps, and write clear, precise descriptions of the actions taking place within each segment.
Your annotations will contribute to datasets used in the development and evaluation of AI systems.
Key Responsibilities
Watch assigned videos and understand the overall activity before annotating.
Segment videos into logical steps or events based on visible action boundaries.
Write clear, concise and grammatically correct descriptions for each segment.
Assign accurate start and end timestamps.
Maintain consistent wording, tense and annotation style across tasks.
Follow provided annotation guidelines, taxonomies and style requirements.
Flag unclear, corrupted or incorrectly labelled videos.
Incorporate reviewer and QA feedback to improve annotation quality.
Meet required productivity and quality standards.
Maintain confidentiality of project data and materials.
Requirements
C1-level English proficiency or higher.
Strong written English, including grammar, vocabulary and sentence construction.
Excellent attention to detail and ability to identify subtle changes in videos.
Ability to describe physical actions clearly and concisely.
Comfortable using web-based tools and annotation software.
Ability to follow detailed instructions and work independently.
Reliable internet connection and access to a computer with a modern browser.
Basic computer and spreadsheet skills.
Preferred Qualifications
Previous experience is helpful but not required. Relevant experience may include:
Data annotation or labelling
Video annotation
Transcription or content moderation
AI or machine-learning projects
Computer vision, NLP or robotics datasets
Writing instructions, SOPs, recipes or how-to content
Annotation tools such as CVAT, Label Studio, Scale or similar platforms
Training on the project's annotation workflow will be provided.
Selection Process
English proficiency assessment (C1 verification)
Annotation skills test involving video segmentation and step writing
Interview or calibration round
Onboarding and guideline training
Compensation
$178 per month
This is a remote opportunity suitable for detail-oriented candidates who can translate visual activities into accurate, step-by-step written instructions.