AI Evaluator – Uber AI (Remote)
Evaluated AI-generated responses for creativity, coherence, empathy, and accuracy across multiple domains. Provided structured feedback to enhance AI performance and improve data quality. Performed quality assurance of narrative flow, conversational tone, and clarity for different content types. • Assessed response quality against psychological and content-development expectations. • Identified issues affecting naturalness, engagement, and correctness. • Contributed to research by flagging effective and compelling responses. • Worked independently in a fully remote setting with deadlines.