AI Evaluation Practice on CrowdGen & Populii platform (2025–Present)
Provided AI response evaluation by assessing generated outputs for accuracy, instruction following, completeness, and safety. Identified factual errors in model responses and compared/ranked responses according to quality criteria. Produced evaluation justifications and completed mock AI evaluator assessments as part of practice. • Reviewed AI-generated text for factual correctness and safety • Checked whether responses complied with given instructions • Assessed completeness against task requirements • Wrote rationale for ratings and rankings