AI Evaluation & Content Analysis — Independent Contractor (Remote)
Independently evaluated AI-generated text responses for factual accuracy, coherence, reasoning quality, and policy compliance across a wide range of prompts. Applied detailed evaluation rubrics to compare and rank multiple model outputs, including handling ambiguous or partially accurate claims with nuanced judgment. Conducted online research and fact verification to validate claims, detect misinformation, and identify hallucinations or inconsistencies. • Assessed instruction adherence, tone alignment, response completeness, and contextual accuracy • Performed comparative response analysis and ranking based on provided criteria • Reviewed for hallucinations, inconsistencies, and misleading information • Maintained consistent, productive work in remote web-based evaluation environments