For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
A

Anuoluwapo A.

Researcher, Annotator and Linguist

Italy flagMilan, Italy

Key Skills

Software

Don't disclose
Other

Top Subject Matter

Multilingual NLP for low-resource languages
Multilingual dataset quality assessment
Evaluation Domain Expertise

Top Data Types

TextText
AudioAudio
DocumentDocument
ImageImage

Top Task Types

Entity (NER) ClassificationEntity (NER) Classification
Audio RecordingAudio Recording
TranscriptionTranscription
Emotion RecognitionEmotion Recognition
SegmentationSegmentation
ClassificationClassification
Object DetectionObject Detection
Text GenerationText Generation
Question AnsweringQuestion Answering
Text SummarizationText Summarization
RLHFRLHF
Fine-tuningFine-tuning
Evaluation/RatingEvaluation/Rating
Data CollectionData Collection
Prompt + Response Writing (SFT)Prompt + Response Writing (SFT)

Freelancer Overview

Researcher and Linguist (voluntary based research) at MASAKHANE Kilifi, Kenya (text annotation and linguistic dataset cu. Brings 9+ years of professional experience across legal operations, contract review, compliance, and structured analysis. Core strengths include Don't disclose and Other. Education includes Master of Science, University of Trento (2026) and Bachelor of Arts, University of Lagos (2021). AI-training focus includes data types such as Text, Audio, and Document and labeling workflows including Entity (NER) Classification, Evaluation, and Rating.

Labeling Experience

Researcher and Linguist (voluntary based research) at MASAKHANE Kilifi, Kenya (text annotation and linguistic dataset curation)

Don't discloseTextTextEntity (NER) ClassificationEntity (NER) Classification

Delivered large-scale annotation coordination for multilingual NLP tasks in low-resource African languages. Led Named Entity Recognition and Parts of Speech tagging projects across distributed annotators and translators to produce model-ready text datasets. Supported downstream evaluation and human quality checks to ensure label consistency and usefulness for training and assessment. • Coordinated 6 annotators on a combined 9,500-text-dataset project • Managed translation and localization workflows for Yoruba and related languages • Performed human model evaluation for news-domain reading comprehension • Annotated 3,000+ samples for Nigerian Pidgin topic classification

2020 - Present

Lead (Linguistic) Data Officer at LELAPA AI, Johannesburg, SA

OtherTextText

Created and operationalized data quality evaluation for multilingual linguistic datasets used in finetuning and model training. Conducted human evaluation of model outputs using automated metrics (WER, CER) and performed error analysis to identify deficiencies for iterative dataset improvements. Enabled higher-quality annotation and training data through refined curation, governance, licensing, and automated pipeline operations. • Designed linguistic human quality metric and co-designed automated quality evaluation • Revamped data achieving/curation/licensing/governance processes • Automated file transfers between AWS S3 and Azure Blob Storage • Performed human evaluation on monolingual and code-switched datasets using WER/CER

2022 - 2026

Lead, Text Generation and Speech Curation (grant) at YORUBANAMES.COM, Lagos, Nigeria

OtherAudioAudioAudio RecordingAudio Recording

Managed the curation and generation of aligned text and audio datasets for text-to-speech and automatic speech recognition. Coordinated voice artists and a large volunteer pool to record speech datasets and ensured the resulting audio aligned with cleaned, curated text. Supported multilingual community data efforts by localizing web and Android content for Yoruba Common Voice. • Led 2 voice artists to record audio for TTS datasets • Managed 80 volunteers for ASR recording workflows • Curated, generated, coordinated, and cleaned aligned text for ASR/TTS • Coordinated Yoruba localization into Mozilla Common Voice via Pontoon

2022 - 2023

Linguist (contract) at ST COMMUNICATIONS, Cape Town, SA

OtherTextText

Worked on a Yoruba-oriented machine translation effort to improve target language quality for Google Translate. Identified problematic source sentences that degrade translation quality and optimized the translation workflow using a machine translation tool. Also supported speech-related dataset recording and evaluation by managing voice contributors and providing contextual training samples. • Identified source sentences hindering translation quality • Optimized translation workflow using Google Transform 3 MT tool • Led 2 voice artists for TTS dataset recording • Provided 50 contextual sample contents for model evaluation and created rule-based grammar for capitalised words

2022 - 2022

Yorùbá Language Phonetician (contract) at TELUS INTERNATIONAL, Vancouver, Canada

Don't discloseAudioAudioTranscriptionTranscription

Provided linguistically grounded preparation and validation of Yoruba speech and text corpora for text-to-speech and speech recognition projects. Authored rule-based style guides and grammar to improve downstream model development and evaluation. Performed evaluation of contextual numeral usage and supplied contextual sample contents as training data. • Led development of rule-based style guides for Yoruba numerical modeling • Supervised corpus validation for grammatical correctness, dialect differences, and audio quality • Evaluated cardinal/ordinal numeral formats in contextual use • Provided 50 contextual sample contents for model evaluation

2022 - 2022

Education

U

University of Trento

Master of Science, Computational Modeling of Language and Cognition

Master of Science
2024 - 2026
U

UAL Creative Computing Institute

Certificate in Conversational AI Interfaces, Introduction to Conversational AI Interfaces

Certificate in Conversational AI Interfaces
2021 - 2021

Work History

L

Lelapa Ai

Lead (Linguistic) Data Officer

Johannesburg
2024 - 2026
M

Masakhane

Researcher and Linguist

Kilifi
2020 - 2026