For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
J
Jerry M.

Jerry M.

AI Response Evaluator / Data & QA Contributor (Afterquery Experts — Remote)

USA flagRaleigh, Usa

Key Skills

Software

Scale AIScale AI

Top Subject Matter

AI language evaluation and RLHF-style response ranking
RLHF response rater work for AI alignment and reliability
Language evaluation and quality checking for AI-related text tasks

Top Data Types

TextText

Top Task Types

RLHFRLHF

Freelancer Overview

AI Response Evaluator / Data & QA Contributor (Afterquery Experts — Remote). Brings 3+ years of professional experience across complex professional workflows, research, and quality-focused execution. Core strengths include Afterquery Experts, Outlier, and RWS. Education includes Bachelor of Science, Mount Kenya University. AI-training focus includes data types such as Text and labeling workflows including Evaluation, Rating, and RLHF.

Labeling Experience

AI Response Evaluator / Data & QA Contributor (Afterquery Experts — Remote)

TextText

Evaluate AI-generated responses for logic, accuracy, clarity, tone, relevance, and instruction-following using rubric-based scoring. Identify hallucinations, unsupported claims, weak reasoning, vague explanations, and contradictions while checking alignment to ideal user intent. Provide structured written justifications to strengthen the human-feedback loop and support AI model improvement. • Logic and factuality assessment of AI outputs • Rubric-based grading and written justifications • Hallucination and contradiction detection • Safety, honesty, helpfulness, and cultural appropriateness review

2026 - Present

AI Data Annotator / RLHF Response Rater (Outlier — Remote)

TextTextRLHFRLHF

Rank and compare AI-generated responses based on helpfulness, factuality, safety, tone, and reasoning quality. Evaluate whether responses follow user instructions and respect constraints while delivering complete and coherent answers. Detect common failure patterns such as overconfidence, hallucinated facts, shallow reasoning, missing context, and poor structure, and communicate decisions clearly. • Response ranking across multiple criteria (helpfulness/safety/tone/reasoning) • Instruction-following and constraint adherence checks • Failure pattern identification (hallucinations, omissions, shallow reasoning) • Written explanations supporting quality ratings

2025 - 2026
Scale AI

Language and AI Evaluation Contributor - RWS

Scale AIScale AITextTextRLHFRLHF

Contributed to language and AI-related projects by reviewing and evaluating text quality against client guidelines. Assessed outputs for grammar, clarity, naturalness, relevance, tone, and cultural appropriateness with an emphasis on consistency. Worked independently on remote evaluation tasks while maintaining accuracy and reliability. • Reviewed content for quality dimensions including clarity and tone • Applied client-specific evaluation guidelines and maintained consistency • Performed remote annotation and review to meet accuracy standards • Supported structured, quality-focused evaluation workflows

2024 - 2025

Language Data / AI Evaluation Contributor (RWS — Remote)

TextText

Review language and AI-related content for grammar, clarity, naturalness, relevance, tone, and cultural appropriateness. Apply client-specific guidelines to ensure consistent evaluation quality across remote annotation tasks. Contribute to structured evaluation work focused on quality assurance and multilingual content awareness. • Text review and quality checking • Rubric/client-guideline compliance • Tone and cultural appropriateness evaluation • Remote independent annotation with reliability standards

2024 - 2025

Education

M

Mount Kenya University

Bachelor of Science, Information Technology

Bachelor of Science
Not specified

Work History

A

Afterquery Experts

AI Response Evaluator and QA Contributor

Raleigh
2026 - Present
R

RWS

Language and AI Evaluation Contributor

Raleigh
2024 - 2025