Senior Multimodal AI Training Specialist, NovaSense AI Ltd.
Designed and maintained multimodal training pipelines across image, text, and audio modalities for production model families. Led dataset curation with QA protocols and inter-annotator agreement scoring to improve label consistency and downstream benchmark performance. Conducted systematic fine-tuning experiments on CLIP/BLIP-2 checkpoints using LoRA/QLoRA and evaluated retrieval accuracy for deployment-ready models. • Managed large-scale vision-language corpus curation (12M samples) across 18 languages • Implemented annotation QA and inter-annotator agreement scoring (81% to 96% consistency) • Fine-tuned CLIP/BLIP-2 with LoRA/QLoRA for cross-modal retrieval (0.87 Recall@5) • Mentored annotation specialists and junior ML engineers; established async review cadence