For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
C
Christian W.

Christian W.

MD/PhD with 7+ years of clinical practice, medical research, and AI-driven

United States Minor Outlying Islands flaglondon, United States Minor Outlying Islands

Key Skills

Software

DoccanoDoccano
LabelboxLabelbox
Label StudioLabel Studio
ProdigyProdigy
Scale AIScale AI
Other

Top Subject Matter

Multi-domain English language content for SFT/RLHF training (science, history, technology, culture)
LLM evaluation and calibration for question answering and summarization
Academic writing quality improvement and AI-assisted composition guidance

Top Data Types

DocumentDocument
Medical DicomMedical Dicom
TextText

Top Task Types

ClassificationClassification
Entity (NER) ClassificationEntity (NER) Classification
Evaluation/RatingEvaluation/Rating
Prompt + Response Writing (SFT)Prompt + Response Writing (SFT)
Question AnsweringQuestion Answering
Text GenerationText Generation
RLHFRLHF

Freelancer Overview

PhD in English Language & Literature with 6+ years of professional writing, editing, and AI language model training at C2 (native-level) proficiency. Expert in prompt-response writing (SFT), rubric-guided evaluation, RLHF feedback, and ranking AI outputs for clarity, accuracy, and style. Crafted 300+ gold-standard responses across science, history, and tech. Published 3 peer-reviewed papers on AI stylistic refinement. Proficient in Google Workspace, LaTeX, Grammarly Pro, and evaluation platforms. Based in London, UK. Seeking English AI Trainer / Data Labeling roles to advance generative AI in professional writing and QA.

Labeling Experience

Freelance AI Language Specialist & Editor - Independent Consultant

TextTextRLHFRLHF

As a Freelance AI Language Specialist and Editor, you produce and refine high-quality English text for AI training datasets and editorial use cases while ensuring factual reliability and stylistic precision. You design and apply evaluation rubrics to rank outputs and support human-in-the-loop quality improvements. The role requires strong grammar and editorial judgement, familiarity with source-based fact-checking, and the ability to work under complex guidelines. • Authored 300+ prompt-response pairs for SFT and RLHF datasets across multiple domains • Edited AI-generated drafts to improve coherence, factual accuracy, and stylistic consistency • Built ranking evaluation rubrics and achieved high inter-rater agreement in blind tests • Conducted fact-checking with primary academic and reference sources to maintain zero factual errors in final outputs

2020 - Present

Freelance AI Language Specialist & Editor — SFT/RLHF dataset authoring and evaluation

OtherTextTextPrompt + Response Writing (SFT)Prompt + Response Writing (SFT)

Authored high-quality prompt-response pairs for SFT and RLHF dataset construction across multiple subject domains, supporting supervised instruction-following and preference learning. Performed iterative editorial refinement of model drafts to improve coherence, factuality, and stylistic consistency before inclusion in training-ready outputs. Conducted fact-checking against primary sources to ensure outputs contained zero factual errors.

2020 - Present
Scale AI

SFT Prompt-Response Writing for LLMs

Scale AIScale AITextTextClassificationClassificationText GenerationText Generation

Authored 300+ high-quality prompt-response pairs for SFT and RLHF datasets across science, history, technology, and culture domains. Used Scale AI, Labelbox, and DeepMind’s internal platform to label, rank, and refine AI outputs. Designed evaluation rubrics achieving 98% inter-rater agreement in blind tests. Improved coherence, factual accuracy, and stylistic consistency by 35%. Fact-checked using JSTOR, OED, Hansard. Trained junior evaluators on AP/Chicago/MLA style adherence. Published 3 peer-reviewed papers on AI stylistic refinement.

2020

Content Evaluator, AI Writing Systems - DeepMind

TextTextRLHFRLHFClassificationClassification

As a Content Evaluator for an AI writing systems team, you evaluated and ranked model-generated responses using structured, multi-dimensional rubrics. You collaborated with ML engineers to refine prompts and system instructions to improve response quality and safety. The work demanded rubric-based analytical skills, clear documentation of evaluation outcomes, and an understanding of hallucination drivers. • Evaluated and ranked AI responses across accuracy, clarity, safety, and voice dimensions • Partnered with ML engineers to iterate on prompt design and system instructions • Produced gold-standard reference answers for calibration in QA and summarization tasks • Trained junior evaluators on style-guide adherence and reasoning transparency

2021 - 2023

Content Evaluator, AI Writing Systems — LLM response ranking and rubric evaluation

TextText

Evaluated and ranked AI-generated responses using multi-dimensional rubrics such as accuracy, clarity, safety, and voice. Collaborated with ML engineers to iterate prompt design and system instructions based on evaluation outcomes to reduce hallucination rates. Produced gold-standard reference answers for model calibration in question answering and summarization tasks.

2021 - 2023

Education

U

University College London (UCL)

PhD, English Language and Literature

PhD
2023 - 2023
U

University College London (UCL)

Doctor of Philosophy, English Language and Literature

Doctor of Philosophy
2019 - 2023

Work History

I

Independent Consultant

Freelance AI Language Specialist & Editor

London
2020 - Present
D

DeepMind

Content Evaluator, AI Writing Systems

London
2021 - 2023