Outlier AI (Scale AI Platform) — AI Training Specialist, Writing & Reasoning (AI trainer and data annotator for RLHF, NLP labels, and code evaluation)
Performed RLHF prompt-response annotation across creative writing, factual Q&A, and logical reasoning, sustaining high quality throughout the evaluation cycle. Evaluated instruction-following, coherence, and factual accuracy using standardized rubrics. Contributed to domain-specific fine-tuning datasets and improved pipeline reliability via edge-case and adversarial prompt flagging. • Completed 400+ RLHF prompt-response tasks with a 96% quality score. • Reviewed and ranked model outputs for rubric-based instruction quality. • Annotated 1,200+ English and Pidgin English text samples for sentiment, NER, and intent. • Reviewed/ranked Python and JavaScript code snippets for correctness, style, and edge-case handling.