AI Content Reviewer / Evaluator (Freelance / Remote) — Scale AI & Independent Contracts
Freelance remote content reviewer and evaluator for conversational and vision-related AI outputs using structured safety and factuality scoring. Assessed model performance using multidimensional rubrics and provided detailed feedback to support iterative RLHF improvements. Identified hallucinations, inconsistencies, errors, and bias within LLM outputs across varied prompts.• Reviewed and rated AI-generated visual and textual responses for safety, factuality, and instruction-following • Scored helpfulness, harmlessness, and honesty using multi-dimensional systems • Detected hallucinations, errors, and bias across a range of prompt types • Supplied structured feedback to model trainers; evaluated 500+ responses weekly within SLAs