LLM Reviewer (Freelance) at Outlier London, UK (10/2024 - Present)
Reviewed LLM outputs across varied prompts and contexts, focusing on accuracy, relevance, coherence, and factual correctness. Performed comparative evaluations of multiple model responses and provided detailed feedback and annotations to improve output quality. Developed evaluation methodologies to enhance AI-generated content trustworthiness and reliability. • Rated responses based on coherence, factual correctness, and alignment with expected standards. • Provided annotated feedback to refine LLM performance and response reliability. • Ensured outputs met accuracy and relevance requirements for diverse tasks and contexts. • Contributed to ongoing AI model optimization and research and development efforts.