Freelance AI Trainer / Data Labeler (Various AI Platforms: Outlier, Remotasks, Appen, Scale AI, etc.)
Evaluated and ranked AI-generated responses using detailed guidelines for large language model training and quality improvement. Performed careful scoring to ensure responses met rubric-based requirements and adhered to instruction-following standards. Maintained high labeling accuracy (above 95%) through consistent review and error checking. • Worked on AI response evaluation and ranking tasks across multiple external platforms • Adapted quickly to new project requirements and evolving guidelines • Applied attention to detail to reduce inconsistencies and guideline deviations • Produced reliable labels to support downstream model training and refinement