For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
S
Samia M.

Samia M.

AI Model Evaluator / LLM Trainer (Contract) — Outlier AI (Remote)

USA flagMinneapolis, Usa

Key Skills

Software

LabelboxLabelbox

Top Subject Matter

LLM response evaluation
factual accuracy
policy & safety compliance

Top Data Types

TextText
AudioAudio
DocumentDocument

Top Task Types

RLHFRLHF
Entity (NER) ClassificationEntity (NER) Classification

Freelancer Overview

AI Model Evaluator / LLM Trainer (Contract) — Outlier AI (Remote). Brings 5+ years of professional experience across legal operations, contract review, compliance, and structured analysis. Core strengths include Labelbox. Education includes Bachelor of Arts, University of Minnesota Twin Cities (2024). AI-training focus includes data types such as Text and Audio and labeling workflows including Evaluation, Rating, and RLHF.

Labeling Experience

Labelbox

AI Response Evaluator - Alignerr

LabelboxLabelboxAudioAudioRLHFRLHFEntity (NER) ClassificationEntity (NER) Classification

You evaluated generative AI responses using detailed rating frameworks and annotation tools. You compared multiple model outputs to identify hallucinations, factual inconsistencies, and policy violations. You contributed to reinforcement learning datasets by providing structured explanations to support evaluation decisions. • Rated response quality and reasoning strength using rating frameworks • Detected hallucinations and safety or policy issues • Compared responses across models for relative ranking • Authored structured written rationales for dataset refinement

2025 - Present
Labelbox

AI Response Evaluator (Contract) — Alignerr (Remote)

LabelboxLabelboxTextTextRLHFRLHF

Assessed generative AI outputs using detailed rating frameworks and annotation tools to determine relative quality and reasoning strength. Identified hallucinations, factual inconsistencies, and policy violations in model responses and produced structured explanations supporting evaluation decisions. Contributed to reinforcement learning datasets used to refine AI system performance. • Compared multiple model responses and selected/graded the best outputs • Detected hallucinations and factual/policy issues in generated text • Wrote structured rationale to document evaluation outcomes • Supported RL dataset creation for system refinement

2025 - Present
Labelbox

AI Model Evaluator / LLM Trainer (Contract) — Outlier AI (Remote)

LabelboxLabelboxTextText

Evaluated large language model outputs using structured rubrics to score accuracy, coherence, helpfulness, and safety. Performed detailed fact-checking and source verification to identify errors and hallucinations, then provided written guidance to improve model behavior. Rewrote and refined prompts and responses, and ranked multiple candidate outputs by reasoning and quality. • Rated model responses against rubric criteria for training and improvement • Conducted factual verification and policy/safety compliance checks • Performed prompt engineering and prompt/response rewriting for better datasets • Compared multiple responses and delivered structured feedback for model iteration

2025 - Present

Education

U

University of Minnesota Twin Cities

Bachelor of Arts, Global Studies

Bachelor of Arts
2021 - 2024

Work History

A

Alignerr

AI Response Evaluator

Remote
2025 - Present
O

Outlier AI

AI Model Evaluator

Remote
2025 - Present