Skip to content
OpenTrain AIFor AI Companies

AI Domain Expert for Model Evaluation

Use your professional expertise to evaluate AI-generated responses, apply detailed rubrics, and provide feedback that improves model behavior. This flexible remote contract offers 20+ hours per week and listed rates of $140 to $200 per hour.

Apply now
OpenTrain AI

Generative AI & RLHF

Remote Hourly · $140–$200/hr

$140–$200/hr

Compensation

230 countries

Eligibility

Entry

Experience

Aug 4, 2026

Posted

Open to applicants in

United States
+ more
  • Åland Islands
  • Albania
  • Algeria
  • American Samoa
  • Andorra
  • Angola
  • Anguilla
  • Antarctica
  • Antigua & Barbuda
  • Argentina
  • Armenia
  • Aruba
  • Australia
  • Austria
  • Azerbaijan
  • Bahamas
  • Bahrain
  • Bangladesh
  • Barbados
  • Belgium
  • Belize
  • Benin
  • Bermuda
  • Bhutan
  • Bolivia
  • Bosnia & Herzegovina
  • Botswana
  • Bouvet Island
  • Brazil
  • British Indian Ocean Territory
  • British Virgin Islands
  • Brunei
  • Bulgaria
  • Burkina Faso
  • Burundi
  • Cambodia
  • Cameroon
  • Canada
  • Cape Verde
  • Caribbean Netherlands
  • Cayman Islands
  • Central African Republic
  • Chad
  • Chile
  • Christmas Island
  • Cocos (Keeling) Islands
  • Colombia
  • Comoros
  • Congo - Brazzaville
  • Cook Islands
  • Costa Rica
  • Côte d’Ivoire
  • Croatia
  • Curaçao
  • Cyprus
  • Czechia
  • Denmark
  • Djibouti
  • Dominica
  • Dominican Republic
  • Ecuador
  • Egypt
  • El Salvador
  • Equatorial Guinea
  • Eritrea
  • Estonia
  • Eswatini
  • Ethiopia
  • Falkland Islands (Islas Malvinas)
  • Faroe Islands
  • Fiji
  • Finland
  • France
  • French Guiana
  • French Polynesia
  • French Southern Territories
  • Gabon
  • Gambia
  • Georgia
  • Germany
  • Ghana
  • Gibraltar
  • Greece
  • Greenland
  • Grenada
  • Guadeloupe
  • Guam
  • Guatemala
  • Guernsey
  • Guinea
  • Guinea-Bissau
  • Guyana
  • Haiti
  • Heard & McDonald Islands
  • Honduras
  • Hungary
  • Iceland
  • India
  • Indonesia
  • Ireland
  • Isle of Man
  • Israel
  • Italy
  • Jamaica
  • Japan
  • Jersey
  • Jordan
  • Kazakhstan
  • Kenya
  • Kiribati
  • Kosovo
  • Kuwait
  • Kyrgyzstan
  • Laos
  • Latvia
  • Lebanon
  • Lesotho
  • Liberia
  • Liechtenstein
  • Lithuania
  • Luxembourg
  • Madagascar
  • Malawi
  • Malaysia
  • Maldives
  • Mali
  • Malta
  • Marshall Islands
  • Martinique
  • Mauritania
  • Mauritius
  • Mayotte
  • Mexico
  • Micronesia
  • Moldova
  • Monaco
  • Mongolia
  • Montenegro
  • Montserrat
  • Morocco
  • Mozambique
  • Namibia
  • Nauru
  • Nepal
  • Netherlands
  • New Caledonia
  • New Zealand
  • Nicaragua
  • Niger
  • Nigeria
  • Niue
  • Norfolk Island
  • North Macedonia
  • Northern Mariana Islands
  • Norway
  • Oman
  • Pakistan
  • Palau
  • Palestine
  • Panama
  • Papua New Guinea
  • Paraguay
  • Peru
  • Philippines
  • Pitcairn Islands
  • Poland
  • Portugal
  • Puerto Rico
  • Qatar
  • Réunion
  • Romania
  • Rwanda
  • Samoa
  • San Marino
  • São Tomé & Príncipe
  • Saudi Arabia
  • Senegal
  • Serbia
  • Seychelles
  • Sierra Leone
  • Singapore
  • Sint Maarten
  • Slovakia
  • Slovenia
  • Solomon Islands
  • South Africa
  • South Georgia & South Sandwich Islands
  • South Korea
  • Spain
  • Sri Lanka
  • St. Barthélemy
  • St. Helena
  • St. Kitts & Nevis
  • St. Lucia
  • St. Martin
  • St. Pierre & Miquelon
  • St. Vincent & Grenadines
  • Suriname
  • Svalbard & Jan Mayen
  • Sweden
  • Switzerland
  • Taiwan
  • Tajikistan
  • Tanzania
  • Thailand
  • Timor-Leste
  • Togo
  • Tokelau
  • Tonga
  • Trinidad & Tobago
  • Tunisia
  • Türkiye
  • Turkmenistan
  • Turks & Caicos Islands
  • Tuvalu
  • U.S. Outlying Islands
  • U.S. Virgin Islands
  • Uganda
  • United Arab Emirates
  • United Kingdom
  • United States
  • Uruguay
  • Uzbekistan
  • Vanuatu
  • Vatican City
  • Vietnam
  • Wallis & Futuna
  • Western Sahara
  • Zambia
  • Zimbabwe

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI recruits contributors for specialized projects where human expertise helps shape how advanced AI systems work, and creating an OpenTrain account is free.

This role is part-time remote contractor work. You can build a lasting AI training portfolio while applying your professional knowledge to practical model evaluation tasks.

  • Work with OpenTrain AI as the hiring and contracting organization.
  • Build experience in a fast-growing field connecting professional expertise with AI development.
  • Manage your AI training career and portfolio through OpenTrain.

About AI Model Evaluation

AI systems learn from carefully reviewed examples, structured judgments, and clear human feedback. In model evaluation work, experts assess whether AI-generated responses are accurate, logical, well-supported, and aligned with professional standards.

Your analysis can help improve model behavior across realistic professional scenarios. Prior AI training experience is helpful, but it is not required when you bring strong real-world domain knowledge and professional writing skills.

  • Review and rate AI-generated responses.
  • Provide evidence-based feedback that supports model improvement.
  • Apply human judgment to complex, domain-specific outputs.

The Role

OpenTrain AI is seeking an AI Domain Expert for Model Evaluation to support AI training work involving model outputs and professional documents. You will use project rubrics, source materials, and your domain knowledge to make consistent, well-reasoned judgments.

The role is suitable for professionals with experience in software engineering, finance, data science, legal work, or another relevant field. Strong written English and the ability to explain nuanced decisions are central to the work.

  • Role: AI Domain Expert for Model Evaluation
  • Work type: Part-time remote contractor
  • Time requirement: 20+ hours per week
  • Data type: Text
  • Languages: English
  • Listed pay range: $140 to $200 USD per hour

What You'll Do

You will evaluate model outputs and documentation against project guidelines, recording clear judgments and the evidence behind them. The work may involve reviewing source materials, annotating data, performing quality reviews, and maintaining consistency across repeated tasks.

You will also design and refine prompts based on realistic professional scenarios. Remote collaboration, careful analysis, thoughtful critique, and ethical judgment are important throughout the project.

  • Evaluate AI-generated outputs using project rubrics.
  • Review source materials and model responses against detailed guidelines.
  • Document accurate, logical, evidence-based judgments.
  • Explain decisions clearly so feedback can improve model behavior.
  • Design and refine prompts based on professional scenarios.
  • Annotate data and conduct quality reviews of AI outputs and documentation.
  • Maintain consistency across repeated review tasks.
  • Collaborate remotely while upholding ethical standards.

Requirements

Candidates must bring relevant professional experience in software engineering, finance, data science, legal work, or another related domain, along with a strong record of producing or reviewing professional documents. The ability to assess AI-generated responses for accuracy, logic, and alignment with domain best practices is essential.

You should be comfortable reviewing or editing complex documents, following detailed project instructions, adapting to evolving requirements, and working independently in a remote environment. Experience with data annotation, prompt engineering, or AI output evaluation is helpful but not essential.

  • Professional expertise in software engineering, finance, data science, legal work, or another relevant domain.
  • Strong written English and professional communication skills.
  • Ability to communicate nuanced analytical feedback clearly.
  • Critical thinking, analytical judgment, and careful attention to detail.
  • Ability to review or edit complex professional documents.
  • Ability to follow detailed instructions and adapt to changing requirements.
  • Ethical judgment for high-stakes review work.
  • Prior data annotation, prompt engineering, or AI evaluation experience is helpful but not required.

Who Should Apply

This opportunity is designed for professionals who want to apply specialized knowledge to cutting-edge AI development. It is marked entry level for AI training, so you do not need previous experience in the industry if you can demonstrate strong domain judgment, professional writing, and careful analysis.

The role is available to candidates in the countries listed for this project and requires English fluency. It may be a strong fit for someone seeking flexible, remote work alongside other professional or personal commitments.

  • Professionals with strong knowledge in a relevant real-world field.
  • Clear, analytical writers who can explain complex judgments.
  • Independent workers who value careful, consistent review.
  • Candidates interested in building an AI training career without requiring prior AI project experience.

How to Apply

Create a free OpenTrain account and apply through OpenTrain AI. Your profile can help you present credible professional experience, discover matching AI training opportunities, and build a portfolio as you contribute to the development of modern AI systems.

  • Apply remotely through OpenTrain AI.
  • Highlight your professional domain expertise and document-review experience.
  • Showcase written communication, analytical judgment, and attention to detail.
  • Begin building a durable portfolio in AI training and data labeling.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Physics Model Evaluation Expert

Use senior physics judgment to evaluate competing AI model solutions, assumptions, and uncertainty. This remote, part-time expert contract pays $80-$160 per hour and requires a physics PhD.

Generative AI & RLHF
Text
Remote · Andorra, United Arab Emirates, Antigua & Barbuda +227 more
English
Part-time · Flexible
Entry level
Hourly · $80–$160/hr

Posted Aug 3, 2026

Finance Model Evaluation Expert

Use your finance expertise to evaluate LLM outputs, identify model weaknesses, and create rubrics and benchmarks for finance-focused AI training. This US contract role offers $100 per hour and requires 20+ hours weekly.

Generative AI & RLHF
Text
Remote · United States
English
Part-time · Flexible
Entry level
Hourly · $100/hr

Posted Jul 16, 2026

AI Model Evaluation Data Scientist

Evaluate and improve AI models through Python development, response ranking, dataset creation, and RLHF on a fully remote, one-month contractor assignment.

Generative AI & RLHF
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 16, 2026