OUTLIER | Spécialiste Évaluation & Entraînement de Modèles IA (LLM / Speech-to-Speech)
Annotation and evaluation of data to improve LLM and Speech-to-Speech model outputs through structured quality and safety assessment. You perform comparative rating of generated responses using criteria such as quality, security, relevance, and fluency, following Side-by-Side/ELO methodologies. You also evaluate Live Speech-to-Speech streams by scoring linguistic accuracy, tone, and naturalness, and write detailed FR/EN justifications to guide algorithm improvements. • Comparative LLM response evaluation (Side-by-Side / ELO Ranking) • Speech-to-Speech live evaluation with scoring of accuracy/tone/naturalness • Creation of detailed analytical rationales in FR/EN • Quality improvement of training data via rigorous, consistent output labeling