AI Data Rater | Appen Inc.
Evaluated AI-generated search results, advertisements, and online content for relevance, accuracy, and cultural appropriateness within large-scale ML projects. Maintained a 98% quality accuracy score while completing 4,500+ evaluation tasks monthly under strict turnaround timelines. Contributed labeled evaluation data used to improve NLP model precision and search ranking performance for enterprise AI clients. • Judged search relevance and content appropriateness against provided criteria • Performed multilingual policy-compliance checks with remote QA support • Applied evolving evaluation guidelines and interpretation of instructions • Ensured SLA adherence through strict quality and productivity targets