Senior Data Labeling Specialist & QA Lead (Remote AI Contractor)
Delivered 500K+ labeled data points monthly across NLP, vision, and multimodal tasks for Fortune 500 AI clients while meeting SLAs of 95%+. Specialized in RLHF annotation by ranking and rating LLM outputs for helpfulness, harmlessness, and honesty to support fine-tuning pipelines. Authored guidelines and edge-case decision trees and performed IAA audits to ensure label consistency and reduce ambiguity. • Ranked and rated LLM responses for RLHF categories (helpfulness, harmlessness, honesty) • Built and maintained annotation guidelines and edge-case decision trees for 10+ project types • Conducted IAA audits using Cohen's Kappa, maintaining scores above 0.90 • Executed systematic QA sampling at 10–20% to catch and correct label errors before delivery