Lead AI Data Specialist - Toptal
Lead cross-functional teams to implement large-scale Reinforcement Learning from Human Feedback (RLHF) workflows for core foundational model training. Drive evaluation and structural validation of multi-turn conversational datasets to improve reasoning, logic patterns, and semantic accuracy. Design Swahili morphosyntax rules, style guides, and taxonomies while performing edge-case analysis and adversarial red-teaming to reduce hallucinations, bias, and validation failures. • Manage remote teams of annotators, developers, and bilingual editors for RLHF execution • Lead dataset evaluation and structural validation to raise performance metrics • Build linguistically grounded Swahili NLP specifications and taxonomic frameworks • Partner with ML engineers to optimize training data pipelines and red-team prompts for safety and quality