AI Data Generalist & Training Specialist
As an AI Data Generalist & Training Specialist, I curated, cleaned, and annotated audio and conversational text datasets for AI speech recognition, NLP, and generative models. I designed annotation guidelines to ensure consistent and accurate labeling across diverse data types and provided structured feedback to support model refinement. My responsibilities included data quality audits, workflow optimization, and cross-functional collaboration to ensure high-fidelity datasets for production AI systems. • Maintained 98%+ annotation accuracy in phonetic alignment, diarization, and contextual tagging. • Developed guidelines for labeling regional accents, dialects, and code-switching. • Conducted output evaluation, bias/hallucination identification, and RLHF support. • Ensured data security and compliance with all confidentiality protocols.