Skip to content
OpenTrain AIFor AI Companies

Linguistics Evaluation Specialist

Review conversations between people and language models, assess instruction-following and language quality, and provide clear feedback. This part-time contract pays $40 to $50 per hour and requires at least 20 hours weekly.

Apply now
OpenTrain AI

Generative AI & RLHF

Remote Hourly · $40–$50/hr

$40–$50/hr

Compensation

230 countries

Eligibility

Entry

Experience

Aug 12, 2026

Posted

Open to applicants in

United States
+ more
  • Åland Islands
  • Albania
  • Algeria
  • American Samoa
  • Andorra
  • Angola
  • Anguilla
  • Antarctica
  • Antigua & Barbuda
  • Argentina
  • Armenia
  • Aruba
  • Australia
  • Austria
  • Azerbaijan
  • Bahamas
  • Bahrain
  • Bangladesh
  • Barbados
  • Belgium
  • Belize
  • Benin
  • Bermuda
  • Bhutan
  • Bolivia
  • Bosnia & Herzegovina
  • Botswana
  • Bouvet Island
  • Brazil
  • British Indian Ocean Territory
  • British Virgin Islands
  • Brunei
  • Bulgaria
  • Burkina Faso
  • Burundi
  • Cambodia
  • Cameroon
  • Canada
  • Cape Verde
  • Caribbean Netherlands
  • Cayman Islands
  • Central African Republic
  • Chad
  • Chile
  • Christmas Island
  • Cocos (Keeling) Islands
  • Colombia
  • Comoros
  • Congo - Brazzaville
  • Cook Islands
  • Costa Rica
  • Côte d’Ivoire
  • Croatia
  • Curaçao
  • Cyprus
  • Czechia
  • Denmark
  • Djibouti
  • Dominica
  • Dominican Republic
  • Ecuador
  • Egypt
  • El Salvador
  • Equatorial Guinea
  • Eritrea
  • Estonia
  • Eswatini
  • Ethiopia
  • Falkland Islands (Islas Malvinas)
  • Faroe Islands
  • Fiji
  • Finland
  • France
  • French Guiana
  • French Polynesia
  • French Southern Territories
  • Gabon
  • Gambia
  • Georgia
  • Germany
  • Ghana
  • Gibraltar
  • Greece
  • Greenland
  • Grenada
  • Guadeloupe
  • Guam
  • Guatemala
  • Guernsey
  • Guinea
  • Guinea-Bissau
  • Guyana
  • Haiti
  • Heard & McDonald Islands
  • Honduras
  • Hungary
  • Iceland
  • India
  • Indonesia
  • Ireland
  • Isle of Man
  • Israel
  • Italy
  • Jamaica
  • Japan
  • Jersey
  • Jordan
  • Kazakhstan
  • Kenya
  • Kiribati
  • Kosovo
  • Kuwait
  • Kyrgyzstan
  • Laos
  • Latvia
  • Lebanon
  • Lesotho
  • Liberia
  • Liechtenstein
  • Lithuania
  • Luxembourg
  • Madagascar
  • Malawi
  • Malaysia
  • Maldives
  • Mali
  • Malta
  • Marshall Islands
  • Martinique
  • Mauritania
  • Mauritius
  • Mayotte
  • Mexico
  • Micronesia
  • Moldova
  • Monaco
  • Mongolia
  • Montenegro
  • Montserrat
  • Morocco
  • Mozambique
  • Namibia
  • Nauru
  • Nepal
  • Netherlands
  • New Caledonia
  • New Zealand
  • Nicaragua
  • Niger
  • Nigeria
  • Niue
  • Norfolk Island
  • North Macedonia
  • Northern Mariana Islands
  • Norway
  • Oman
  • Pakistan
  • Palau
  • Palestine
  • Panama
  • Papua New Guinea
  • Paraguay
  • Peru
  • Philippines
  • Pitcairn Islands
  • Poland
  • Portugal
  • Puerto Rico
  • Qatar
  • Réunion
  • Romania
  • Rwanda
  • Samoa
  • San Marino
  • São Tomé & Príncipe
  • Saudi Arabia
  • Senegal
  • Serbia
  • Seychelles
  • Sierra Leone
  • Singapore
  • Sint Maarten
  • Slovakia
  • Slovenia
  • Solomon Islands
  • South Africa
  • South Georgia & South Sandwich Islands
  • South Korea
  • Spain
  • Sri Lanka
  • St. Barthélemy
  • St. Helena
  • St. Kitts & Nevis
  • St. Lucia
  • St. Martin
  • St. Pierre & Miquelon
  • St. Vincent & Grenadines
  • Suriname
  • Svalbard & Jan Mayen
  • Sweden
  • Switzerland
  • Taiwan
  • Tajikistan
  • Tanzania
  • Thailand
  • Timor-Leste
  • Togo
  • Tokelau
  • Tonga
  • Trinidad & Tobago
  • Tunisia
  • Türkiye
  • Turkmenistan
  • Turks & Caicos Islands
  • Tuvalu
  • U.S. Outlying Islands
  • U.S. Virgin Islands
  • Uganda
  • United Arab Emirates
  • United Kingdom
  • United States
  • Uruguay
  • Uzbekistan
  • Vanuatu
  • Vatican City
  • Vietnam
  • Wallis & Futuna
  • Western Sahara
  • Zambia
  • Zimbabwe

The work

You will review conversations between people and large language models. Your evaluations will help improve how AI systems understand language, follow instructions, and respond in context.

You will use written guidelines to assess model responses and record clear findings. You will also work with project coordinators to clarify unclear instructions or evaluation cases.

  • Review conversation transcripts for accuracy, clarity, and completeness.
  • Assess whether language model responses follow the provided instructions and guidelines.
  • Find subtle grammar, language-use, and context errors.
  • Give detailed written and verbal feedback on response quality.
  • Write organized reports with findings and recommendations.
  • Help clarify evaluation guidelines when they are ambiguous.

What it pays and takes

This is a part-time contractor role for a language model evaluation project. No previous AI experience is required, but strong linguistics knowledge and careful language analysis are essential.

  • Pay: $40 to $50 USD per hour.
  • Time: At least 20 hours per week.
  • Location: You must be based in one of the listed eligible countries.
  • Language: Native proficiency in U.S. English.
  • Background: Advanced academic training or extensive professional experience in linguistics, language analysis, or a related field.
  • Strengths: Excellent written communication, nuanced language judgment, and close attention to detail.
  • Availability: Able to begin contributing immediately upon selection.
  • Preferred: Experience evaluating work against guidelines or rubrics, plus the ability to provide written and verbal insights.

How it works

Apply on OpenTrain with your resume and then complete the application on the hiring site.

About AI training work

OpenTrain AI is the hiring and contracting organization for this role. AI training work uses human reviews, examples, and feedback to help artificial intelligence systems produce more accurate and useful results, and linguistics specialists are paid for the language expertise needed to judge subtle differences in meaning and quality.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

LLM Evaluation Specialist

Evaluate large language models and help improve their performance in a remote, part-time freelance role. Open to students and graduates in any academic field, with pay up to $40 per hour.

Generative AI & RLHF
Text
Remote · United States
English
Part-time · Flexible
Entry level
Hourly · $40/hr

Posted Sep 30, 2026

Malay Language LLM Evaluation Analyst

Review Malay and English language model responses, validate claims, solve reasoning tasks, and write detailed feedback. This remote contractor role requires 20+ hours per week.

Generative AI & RLHF
Text
Remote · Worldwide
Malay, English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Lithuanian Language AI Evaluation Expert

Review and rate AI-generated Lithuanian content for accuracy, fluency, cultural fit, safety, and meaning. This remote contractor role pays up to $25 per hour and requires 20 or more hours each week.

Generative AI & RLHF
Text
Remote · Lithuania
Lithuanian, English
Part-time · Flexible
Entry level
Hourly · $25/hr

Posted Aug 19, 2026