AI Data Annotator / Tasker | Outlier (Scale AI)
Served as an AI Data Annotator at Outlier (Scale AI), specializing in the annotation and evaluation of large language models' outputs. Conducted Reinforcement Learning from Human Feedback (RLHF) tasks to assess and improve LLM responses. Performed text data labeling across multiple domains such as code, science, mathematics, and general knowledge. • Evaluated and ranked AI-generated text for quality, coherence, and factual accuracy. • Annotated model outputs, flagged inconsistencies, and identified edge cases. • Maintained high-quality standards by adhering to complex guidelines and rubrics. • Provided structured feedback to enhance model performance and handled sensitive content moderation.