Multimodal Data Annotator (English & Malayalam) at micro1 (2026)
Executed high-precision acoustic transcription and audio metadata logging for bilingual (English & Malayalam) video datasets. Documented multiple auditory elements including speech, background score, ambient environmental noise, and sound effects with precise time-stamps and quality metrics. Ensured richly annotated audio tracks to train computer vision and multimodal AI models using time-aligned auditory context. • High-precision acoustic transcription (English & Malayalam) • Time-stamped audio metadata logging for video segments • Labeled speech, background score, ambient noise, and sound effects • Added quality metrics to support multimodal training pipelines