AI Content Evaluator & Data Annotator (Remote Freelance Platform)
Served as an AI content evaluator and data annotator performing comparative assessments of AI-generated outputs to determine superiority. Applied detailed guidelines to rank and optimize multimodal responses using factual accuracy, logical consistency, and safety alignment as primary criteria. Provided descriptive feedback on edge cases and observed model failures to support engineering improvements and reduce bias and hallucinations. • Performed 30+ specialized validation tasks across complex text, code, and multimodal model responses • Labeled and annotated large-scale datasets according to project guidelines to create clean training sets • Isolated nuances and edge cases and reported findings for algorithmic bias and hallucination reduction • Selected and justified better responses using evaluation and rating rubric logic