AI Trainer at Prolific Remote (2025–Present)
Provided quality evaluation of large language model responses for accuracy, relevance, clarity, and helpfulness. Conducted fact-checking and verification to identify errors or unsupported claims. Compared multiple AI-generated outputs and selected the best response based on project criteria. • Rated responses against rubric criteria • Performed response evaluation and accuracy checks • Created prompts to test/improve model performance • Contributed to development and improvement of LLMs