AI Training Specialist / Data Annotator
Conducted high-volume evaluation, ranking, and response generation tasks to train conversational Large Language Models (LLMs). The scope involved reviewing side-by-side model outputs to grade them on accuracy, tone consistency, formatting rules, and helpfulness. Main tasks included writing complex, adversarial prompts to test the limits of AI reasoning, rewriting subpar model responses, and providing comprehensive written rationales to justify evaluations. Maintained exceptional quality standards by adhering strictly to multi-page, complex project rubrics and cross-referencing external facts to ensure data truthfulness.