Freelance AI Trainer & RLHF Specialist (DataAnnotation.tech / Outlier AI / Scale AI)
Completed RLHF preference-ranking tasks for large language model fine-tuning with a focus on evidence-based judgments. Evaluated factual accuracy, logical coherence, and tone of LLM outputs across scientific, statistical, and public-health domains while flagging hallucinations. Authored prompt–response pairs for instruction fine-tuning and participated in calibration to resolve edge-case disagreements. • Completed 1,800+ RLHF preference-ranking tasks • Achieved >95% inter-annotator agreement against gold-standard labels • Delivered 10,000+ labeled data units across text, image, and structured formats for LLM/search-quality systems • Wrote prompt–response pairs covering data analysis, epidemiology, and research methodology