AI Content Evaluator / Senior Data Annotator | Outlier
Serve as an AI content evaluator ranking AI-generated text using Reinforcement Learning from Human Feedback (RLHF) methods. Perform systematic review to assess response quality with emphasis on factuality and safety. Support model improvement by identifying problematic outputs and the linguistic issues that lead to hallucinations or biased behavior.• Rank AI-generated text outputs via RLHF-based evaluation criteria.• Conduct factuality audits to flag hallucinations and incorrect claims.• Run safety checks to identify unsafe or bias-inducing content patterns.• Provide evaluation findings that help improve LLM accuracy, factuality, and safety.