AI Data Annotation & Evaluation Specialist (Outlier AI | Remote)
Worked on annotating and evaluating AI-generated responses for quality, accuracy, and compliance using evaluation frameworks. Graded outputs and ensured consistent human preference judgments across evaluation tasks. Contributed to structured labeling workflows supporting large-scale dataset needs. • Reviewed LLM responses against quality and compliance criteria • Applied evaluation frameworks for response grading • Maintained high consistency in human preference/RLHF-style tasks • Built and followed structured labeling workflows