AI Evaluation Contributor - Outlier AI
Contributed to AI model evaluation and data preparation workflows across Outlier AI platforms, focusing on prompt evaluation and response ranking. Applied critical thinking to judge outputs against quality rubrics and inform downstream improvements. Worked in a structured human-in-the-loop setup to guide, correct, and validate automated AI agent outputs. • Performed prompt and response evaluation against predefined rubrics • Ranked and validated AI outputs to support quality fine-tuning • Collaborated within human-in-the-loop processes • Ensured consistency and quality of model outputs