AI Content Evaluator (Generalist)
This role focuses on evaluating and rating Large Language Model (LLM) outputs for accuracy and logical consistency. The specialist reviews AI-generated text, identifies hallucinations or formatting issues, and documents safety concerns. Continuous, detailed feedback is provided to refine the AI model's helpfulness and reliability. • Performed accuracy assessments on AI-generated responses • Identified and documented hallucinations and safety violations • Provided detailed structured feedback to improve model training • Ensured content alignment with factual and logical standards.