AI Language & Response Evaluation Contributor | Project-Based (Remote / Nairobi, Kenya)
Reviewed AI-generated responses for instruction following, factual accuracy, clarity, quality, and realism as part of an AI Language & Response Evaluation project. Provided concise feedback by identifying hallucinations, formatting issues, and weak reasoning. Applied rubrics to assess model output quality and determine pass/fail or improvement notes. • Tested responses against prompts and evaluation criteria • Flagged hallucinated or incorrect statements • Evaluated formatting consistency and reasoning strength • Produced actionable written feedback for model improvement