Senior Image/Video Annotator – Multimodal AI Training (Appen, remote, US)
Senior Image/Video Annotator for multimodal AI training specializing in emotion and visual cue labeling in short video clips. Annotated facial/body emotion categories and mapped ambient mood, lighting, and spatial relationships to support downstream multimodal understanding. Ensured consistency with an 80+ page annotation stylebook and ongoing calibration and QA processes. • Labeled emotions (e.g., happy, angry, confused, neutral) and intensity/visual cues from video frames. • Annotated mood/lighting conditions (bright, dim, backlit; suspenseful/cheerful/somber) and spatial relations (behind/beside/above/approaching). • Produced frame-accurate descriptions of scene transitions and character interactions. • Achieved 96% consistency with gold labels and maintained a 98% QA score for 8 consecutive months.