Skip to content
OpenTrain AIFor AI Companies

AI Safety Model Evaluator

Evaluate AI model responses for safety, factual accuracy, policy compliance, and alignment across high-risk topics. This remote expert contract pays $60-$70 per hour and requires 20+ hours per week.

OpenTrain AI

Generative AI & RLHF

Remote Hourly · $60–$70/hr

$60–$70/hr

Compensation

33 countries

Eligibility

Expert

Experience

Jul 22, 2026

Posted

Open to applicants in

United States Denmark Estonia Finland Ireland Latvia Lithuania Norway Sweden Austria Belgium France Germany Netherlands Switzerland United Kingdom Albania Bosnia & Herzegovina Croatia Greece Italy Malta Portugal Serbia Slovenia Spain Bulgaria Czechia Hungary Moldova Poland Romania Slovakia

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help people discover specialized projects, build a professional AI training profile, and apply in minutes. Creating an OpenTrain account is free.

  • Remote contract work with an expert-level AI training focus
  • A profile that helps you document relevant experience and grow your AI training career
  • Access to opportunities shaping how advanced AI systems behave

About AI Safety Model Evaluation

AI training is the human side of building artificial intelligence. Evaluators review model responses and provide structured judgments that help AI systems become more accurate, useful, safe, and aligned. This work is part of a fast-growing technology field spanning RLHF, supervised fine-tuning, benchmarking, and safety evaluation.

  • Assess examples produced by modern AI systems
  • Identify unsafe, inaccurate, misleading, or policy-violating behavior
  • Help improve model performance through consistent, evidence-based feedback

The Role

OpenTrain is recruiting an AI Safety Model Evaluator to assess the safety, quality, factual accuracy, and alignment of frontier AI model responses. You will apply structured standards to complex, policy-sensitive, and ambiguous topics, helping improve model behavior and safety performance.

The work spans high-risk content areas including misinformation, political persuasion, self-harm, violence, cyber, and biosecurity. Success requires excellent written English, strong critical thinking, analytical reasoning, and consistent judgment.

  • Role: AI Safety Model Evaluator
  • Work type: Remote contractor and part-time opportunity
  • Experience level: Expert
  • Pay: $60-$70 per hour
  • Time requirement: 20+ hours per week, with a default commitment of 40 hours per week
  • Language: English
  • Eligible locations: United States and selected countries across Europe

What You'll Do

You will evaluate AI-generated content against safety, factual accuracy, policy compliance, and overall quality standards. Your reviews will support model alignment, safety benchmarking, and the development of stronger evaluation practices.

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and quality
  • Review nuanced content across sensitive and high-impact domains
  • Apply and refine evaluation rubrics for RLHF, supervised fine-tuning, and AI safety benchmarking
  • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations
  • Provide structured feedback that supports stronger model alignment and safety performance
  • Collaborate with AI researchers and safety teams on evaluation initiatives

Required Qualifications

Applicants must bring substantial professional experience and the ability to make careful, consistent judgments in nuanced and policy-sensitive scenarios. A bachelor's degree or higher is required in journalism, communications, psychology, sociology, public policy, law, biology, chemistry, computer science, or a related discipline.

  • Bachelor's degree or higher in a listed or related discipline
  • At least five years of professional experience in AI safety, trust and safety, journalism, public policy, scientific research, security, or a related field
  • Excellent written English
  • Strong critical-thinking and analytical-reasoning skills
  • Ability to evaluate misinformation, self-harm, violence, cyber, biosecurity, and other policy-sensitive content
  • Ability to identify unsafe outputs, hallucinations, reasoning failures, and policy violations
  • Ability to apply evaluation standards consistently

Helpful Background

The following experience is helpful but is presented as preferred background rather than a required qualification. Familiarity with AI safety and evaluation methods can help you contribute effectively to complex review initiatives.

  • Experience with AI safety, RLHF, supervised fine-tuning, trust and safety, or AI evaluation
  • Familiarity with safety policies or content moderation
  • Experience developing or applying evaluation rubrics
  • Experience reviewing complex, high-risk, or ambiguous content

Schedule, Location, and Pay

This is remote contract work available to contractors residing in the United States, Denmark, Estonia, Finland, Iceland, Ireland, Latvia, Lithuania, Norway, Sweden, Austria, Belgium, France, Germany, Liechtenstein, Luxembourg, Monaco, the Netherlands, Switzerland, the United Kingdom, Albania, Bosnia and Herzegovina, Croatia, Greece, Italy, Kosovo, Malta, North Macedonia, Portugal, San Marino, Serbia, Slovenia, Spain, Bulgaria, the Czech Republic, Hungary, Moldova, Poland, Romania, or Slovakia.

  • Hourly rate: $60-$70 USD
  • Minimum expected availability: 20+ hours per week
  • Default commitment: 40 hours per week
  • Contractor engagement
  • Remote work from an eligible country

Build Your AI Training Career With OpenTrain

AI safety evaluation lets experienced professionals directly influence how advanced models respond to difficult and consequential situations. Through OpenTrain, you can build a durable portfolio of AI training experience, show relevant expertise, and find projects aligned with your skills.

  • Create a free OpenTrain account
  • Build a profile highlighting your professional background
  • Apply to eligible AI training opportunities in minutes
  • Grow a portfolio in a rapidly developing AI industry

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Korean AI Safety Model Evaluator

Use Korean fluency, English writing, and cultural judgment to evaluate sensitive AI model behavior through prompt creation, classification, and red-teaming. Remote contractor work is approximately 7 hours weekly at $48-$52 per hour.

Generative AI & RLHF
Text
Remote · Worldwide
Korean, English
Part-time · Flexible
Entry level
Hourly · $48–$52/hr

Posted Sep 4, 2026

Norwegian AI Safety Model Evaluator

Help evaluate and strengthen AI model safety in Norwegian through prompt writing, conversation classification, and adversarial pattern detection. This remote contractor role pays $58 to $62 per hour for about 7 hours weekly.

Generative AI & RLHF
Text
Remote · Worldwide
Norwegian, English
Part-time · Flexible
Entry level
Hourly · $58–$62/hr

Posted Sep 4, 2026

AI Safety LLM Evaluator, French and English

Work remotely as a French and English AI Safety LLM Evaluator, reviewing model responses, red-teaming safety boundaries, and creating evaluation data. Earn $24 to $36 per hour while helping improve safer AI systems.

Generative AI & RLHF
Text
Remote · Worldwide
French
Part-time · Flexible
Intermediate level
Hourly · $24–$36/hr

Posted Apr 3, 2026