About this opportunity
Job Description
Turing is hiring detail-oriented professionals with strong visual interpretation skills to support the training of advanced multimodal AI systems.
You’ll analyze and annotate videos, images, or both, identifying people, objects, environments, actions, relationships, emotions, and storytelling elements. You may also evaluate AI-generated outputs to determine whether they accurately understand the visual content.
Previous image or video annotation experience is preferred but not required.
What You’ll Do
- Analyze short videos and images to identify key visual and narrative elements.
- Annotate subjects, objects, environments, actions, interactions, and emotions.
- Describe sequences of events, scene transitions, and interactions in video content.
- Label visual details such as mood, lighting, and spatial relationships.
- Follow detailed annotation guidelines to maintain consistency and data quality.
- Review AI-generated outputs for accuracy, structure, and relevance.
- Adapt to evolving annotation standards and quality requirements.
- Contribute structured data that helps AI models better understand images, scenes, motion, and context.
Requirements
- Strong English reading, writing, and communication skills.
- Excellent attention to detail.
- Ability to follow detailed and nuanced instructions accurately.
- Comfortable using computers, web-based tools, and handling file uploads.
- Strong analytical skills and ability to interpret context.
- Ability to work independently in a remote environment.
- Reliable desktop or laptop computer.
- Stable internet connection.
- Availability for 20 hours per week, with the required PST overlap.
Preferred Experience
Previous experience with image annotation, video annotation, data labeling, AI training, or visual content evaluation is an advantage but is not mandatory.