Skip to content
OpenTrain AIFor AI Companies

History Domain Reviewer for AI Model Evaluation

Apply advanced historical knowledge to review prompts and AI-generated work for factual accuracy, reasoning quality, and guideline compliance in a remote contract supporting language-model evaluation.

OpenTrain AI

Generative AI & RLHF

100% Remote

Worldwide

Eligibility

Entry

Experience

Aug 15, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We connect contributors with meaningful projects, help them build a professional AI-training profile, and make it easy to apply and grow in a fast-moving industry.

About AI Training Work

AI training is the human side of developing modern artificial intelligence. Expert reviewers assess examples, prompts, and model responses so AI systems can become more accurate, reliable, and useful. This remote work lets specialists apply their knowledge directly to cutting-edge model development.

The Role

OpenTrain is hiring a History Domain Reviewer to support quality assurance for large language model evaluation projects. You will review historical prompts and completed tasks, verify their quality, and provide evidence-based feedback that improves the reliability of AI evaluation data.

This is an opportunity for a historian or related subject-matter expert to apply deep knowledge of historical periods, events, civilizations, and cultural developments to frontier AI quality work.

  • Contract assignment lasting 8 weeks
  • Contractor and part-time engagement
  • Remote work available worldwide
  • English-language role

What You'll Do

You will evaluate historical content across a broad range of eras and topics, using research, analytical judgment, and project guidelines to identify issues and improve task quality.

  • Review and validate prompts covering ancient, medieval, modern, and contemporary history.
  • Evaluate completed tasks for factual accuracy, reasoning quality, consistency, and compliance with guidelines.
  • Identify inaccuracies, logical inconsistencies, hallucinations, and outdated information.
  • Check that prompts are challenging, relevant, and aligned with project objectives.
  • Provide clear, evidence-based feedback to contributors.
  • Escalate ambiguous cases and document findings.
  • Collaborate with project and AI teams to refine evaluation processes.

Requirements

This role requires strong historical knowledge and the ability to assess written material carefully and objectively. A master's degree or higher is required, with study in a history-related discipline preferred.

  • Master's degree or higher in any field
  • History, Archaeology, Classics, Cultural Studies, Anthropology, or a related discipline preferred
  • At least 3 years of professional experience in historical research, academia, education, or a related field
  • Strong knowledge of major historical periods, events, civilizations, and cultural developments
  • Excellent written English
  • Strong research and analytical skills

Schedule and Commitment

The assignment requires roughly 40 hours per week, with at least 4 hours of overlap with Pacific Time. The role is listed with a commitment of 20 or more hours per week, so applicants should be prepared to meet the project's expected workload and overlap requirement.

  • Expected workload: approximately 40 hours per week
  • Minimum listed commitment: 20+ hours per week
  • At least 4 hours of Pacific Time overlap required
  • Contract duration: 8 weeks

Why Join OpenTrain

OpenTrain helps people build careers at the intersection of human expertise and artificial intelligence. As a contractor, you can apply your historical knowledge to meaningful AI evaluation work while maintaining the flexibility of freelance work within the required commitment.

  • Contribute to the quality and reliability of language-model evaluation data
  • Apply specialized historical expertise to advanced AI projects
  • Work remotely with a flexible freelance structure
  • Build an AI-training portfolio around your professional knowledge

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all jobs

AI Domain Expert for Model Evaluation

Evaluate AI-generated responses, apply expert judgment, and provide feedback that improves model behavior. This flexible, worldwide contractor role offers 20+ hours per week and pays $140–$200 USD per hour.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $140–$200/hr

Posted Aug 4, 2026

AI Model Reviewer, Evaluation & RLHF Specialist

Join OpenTrain as an AI Model Reviewer to evaluate and generate high-quality examples, prompts, and rationales that improve model reasoning — remote, contractor work at $100–$180/hr for 20+ hours/week. Ideal for experts in law, education, engineering, science, or writing.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $100–$180/hr

Posted Jul 30, 2026

TV & Movies Domain AI Review Evaluator

Review AI-generated entertainment content for factual accuracy, reasoning quality, hallucinations, and guideline compliance. Use your professional knowledge of television and film in a remote, eight-week US contract.

Generative AI & RLHF
Text
Remote · United States
English
Part-time · Flexible
Entry level

Posted Aug 4, 2026