Senior Data Annotation Specialist — Maximus
Performed RLHF preference-ranking annotation for a Tier-1 LLM client by generating preference outputs from prompt-response pairs. Maintained high-quality standards, including zero rework requests from the client QA team and 100% inter-annotator agreement for NER tasks over a 90-day sprint. Improved team efficiency and model safety by authoring detailed multilingual guidelines and flagging systematic labeling errors early before delivery. • Led a team of 12 annotators on a large-scale RLHF project (14,000+ prompt-response pairs). • Maintained 100% inter-annotator agreement benchmark for Named Entity Recognition tasks. • Authored a 40-page multilingual sentiment annotation guideline document covering English, Swahili, and Luganda. • Identified a label-bleed error in a 6,000-sample medical NLP dataset prior to delivery.