For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
W
Willy N.

Willy N.

AI Training Specialist and Multimodal Data Evaluator, Educator, and Legal and Business Research

Kenya flagNairobi, Kenya

Key Skills

Software

AppenAppen
ClickworkerClickworker
CloudFactoryCloudFactory
CVATCVAT
LionbridgeLionbridge
RemotasksRemotasks
SamaSama
TolokaToloka
TelusTelus

Top Subject Matter

Business Alalysist
Regulatory Compliance & Risk Analysis
Education

Top Data Types

TextText
ImageImage
DocumentDocument

Top Task Types

Bounding BoxBounding Box
PolygonPolygon
SegmentationSegmentation
ClassificationClassification
Entity (NER) ClassificationEntity (NER) Classification
Point/Key PointPoint/Key Point
PolylinePolyline
CuboidCuboid
Object DetectionObject Detection
Text GenerationText Generation
Question AnsweringQuestion Answering
Text SummarizationText Summarization
Evaluation/RatingEvaluation/Rating
Data CollectionData Collection

Freelancer Overview

I have extensive experience in AI training and data labeling, having worked across major industry platforms including Turing, Remotask, CloudFactory, Toloka, Appen, and Telus. My background is defined by high-level technical versatility, ranging from 3D LiDAR annotation and image annotation to large-scale data collection and photo uploads. I am also highly skilled in content moderation, with significant experience in NSFW image evaluation, ensuring that data meets strict safety and quality standards. What sets me apart is my ability to handle complex, multi-modal projects with precision. Whether it is evaluating AI-generated video reconstruction for consistency or providing detailed justifications in AI evaluation tasks, I bring a logical and results-oriented approach to every project. My deep experience across these diverse platforms allows me to quickly adapt to new rubrics and deliver the high-quality data necessary for refining sophisticated AI models.

Labeling Experience

LLM Prompt Evaluation and AI Trainer

TextTextEvaluation/RatingEvaluation/Rating

The project focused on the large-scale evaluation and refinement of Large Language Model (LLM) responses. The primary objective was to improve the model's ability to generate accurate, helpful, and human-like text by subjecting it to rigorous human-in-the-loop testing. This involved analyzing how the AI interpreted various user prompts and ensuring its outputs aligned with specific safety and quality benchmarks. I performed prompt evaluation and comparative analysis of AI-generated responses. My tasks included scoring model outputs based on relevance, factual accuracy, and linguistic naturalness. I was responsible for identifying and logging technical flaws such as hallucinated information, unnatural phrasing, and tone inconsistencies. Additionally, I provided detailed, evidence-based justifications for my ratings to help developers understand the specific "why" behind a model’s success or failure. This was a high-volume, continuous evaluation project spanning two months. I handled hundreds of unique prompt-response pairs, contributing to a vast dataset used to fine-tune the model's conversational capabilities and technical reliability. I maintained high data integrity by adhering to strict Inter-Annotator Agreement (IAA) standards and Gold Standard benchmarks. My work was subject to regular quality audits to ensure consistency with the project's complex rubrics. I focused on qualitative precision, moving beyond simple binary choices to provide nuanced feedback loops that directly informed the model's iterative training process.

2026 - 2026

AI Training and Data Labeling Specialist

ImageImageBounding BoxBounding Box

I have built an extensive background in AI training and data labeling from April 2021 to March 2026, working on various short-term contracts with industry leaders such as CloudFactory, Appen, Telus, Remotask, and Toloka. My experience spans the full spectrum of Computer Vision and Natural Language Processing. In the realm of image annotation, I have mastered technical tasks including Polygon and Segmentation for precise boundary definition, Cuboid and Object Detection for 3D spatial awareness, and the use of Key Points, Polylines, and Point annotation for navigational and tracking AI. I am also highly skilled in Classification, ensuring vast datasets are categorized with high accuracy for improved model recognition. In the field of Generative AI and NLP, I have contributed significantly to model safety and performance through RLHF, Red Teaming, and Fine-Tuning. My work involves high-level Text Generation, Text Summarization, and Question Answering, as well as Entity NER Classification to improve information extraction. Additionally, I have performed high-accuracy Transcription and Evaluation Rating to ensure AI outputs meet strict quality benchmarks. Over these five years, I have demonstrated a unique ability to adapt to diverse project rubrics, consistently delivering the high-fidelity data required to refine sophisticated, human-centric AI models.

2021 - 2026

Education

A

ALX

Cerificate, AI Essentials

Cerificate
2026 - 2026
K

Kasneb

Certified Public Accountant, Accounting

Certified Public Accountant
2026 - 2026

Work History

P

Petanns Driving School & Computer College

Theory Trainer

Kiambu
2026 - Present
N

National Youth Service

Economics & Business Law Trainer

Nairobi
2025 - 2026