Skip to content
OpenTrain AIFor AI Companies

Remote AI training and data labeling jobs in Mexico

Find remote AI training and data labeling jobs open to applicants in Mexico.

  • 100% remote
  • Flexible hours
  • Hourly / per-task pay
  • 257 open roles

Dutch-Speaking LLM Analyst

Work remotely as a Dutch and English LLM Analyst, answering analytical questions, checking claims, and giving detailed feedback to improve AI models. This part-time contractor role requires 20 or more hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
Dutch, English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Arabic-Speaking AI Analyst

Review Arabic and English content, test language model reasoning, validate answers, and write detailed feedback. This remote contractor role requires 20+ hours weekly and strong Arabic and English skills.

Generative AI & RLHF
Text
Remote · Worldwide
Arabic, English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Italian-Speaking LLM Evaluation Analyst

Review Italian and English content for large language models, validate claims, answer reasoning questions, and write clear feedback. This remote contractor role offers flexible hours with a 40-hour weekly commitment and Italian-English fluency required.

Generative AI & RLHF
Text
Remote · Worldwide
Italian, English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

German-Speaking Multilingual LLM Evaluation Analyst

Evaluate written content and analytical questions in German and English to help improve large language models. Research claims, check answers, create training scenarios, and write clear feedback remotely for 20+ hours per week.

Generative AI & RLHF
Text
Remote · Worldwide
German, English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

French-Speaking AI Model Evaluator

Evaluate French-language AI responses through research, reasoning, summaries, and detailed feedback. This remote contractor role requires English and French, a reliable computer, and 20 or more hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
French, English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Chinese-Speaking LLM Content Evaluator

Evaluate Chinese and English content for large language models by checking claims, solving analytical questions, and writing clear feedback. This remote contractor role requires at least 20 hours per week and flexible overlap with US Pacific time.

Generative AI & RLHF
Text
Remote · Worldwide
English, Chinese
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Image and Video Annotation Specialist

Review images and short videos to label subjects, actions, emotions, relationships, and scene changes for AI training. This entry-level contract role is open worldwide and requires English and 20+ hours per week.

Image & Video Annotation
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Go Code Review Developer

Review and improve AI-generated Go code for correctness, efficiency, and reliability. This worldwide contractor role requires 20+ hours per week, strong Go skills, and at least 3 years of software engineering experience.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Physics LLM Evaluation Expert

Create difficult physics problems, detailed solutions, and evaluation benchmarks for large language models. This expert contract role is open worldwide, requires strong physics and graduate-level STEM experience, and offers 20+ hours per week.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Expert level

Posted Jul 17, 2026

Senior LLM Code Evaluation Engineer

Evaluate how language models work with real software repositories, from GitHub issue triage to local testing and bug-fixing tasks. This contract role needs strong Go skills, 3+ years of engineering experience, and 20+ hours each week.

Coding & Software
Computer Code Programming
Remote · India, Pakistan, Nigeria +6 more
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

C++ LLM Evaluation Software Engineer

Evaluate how well large language models solve real C++ software problems. Build code tasks from open-source projects, test repositories locally, and work remotely for 20 or more hours each week.

Coding & Software
Computer Code Programming
Remote · India, Pakistan, Nigeria +6 more
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Biology LLM Evaluation Expert

Help evaluate advanced Biology language models by creating difficult problems, solving them carefully, and writing clear feedback. This part-time contract role is open worldwide to English-speaking Biology experts.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Chemistry AI Training Expert

Create, solve, and explain advanced chemistry problems used to train AI models. This remote contractor role requires deep chemistry knowledge, clear reasoning, and at least 20 hours per week.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Physics Expert for LLM Evaluation

Design and solve difficult physics problems that test large language models, then write clear solutions and feedback. This remote contractor role needs strong physics knowledge, excellent English, and at least 20 hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

LLM Java AI Training Developer

Build Java backend components and help evaluate, fine-tune, and improve large language models. This remote contractor role requires 20+ hours per week, strong English, and solid software development skills.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

JavaScript and TypeScript AI Model Evaluation Developer

Build JavaScript and TypeScript code, evaluate AI model responses, create fine-tuning datasets, and support human feedback training. This remote contractor role runs for one month with 20, 30, or 40 hours per week.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

AI Coding Evaluation And Training Developer

Help train and evaluate AI coding models with Python, JavaScript or TypeScript, response ranking, dataset creation, and RLHF. This worldwide contractor role requires 20+ hours per week and strong English.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Board Game AI Reasoning Evaluator

Create and review board game reasoning tasks, assess AI answers, and build prompts and rubrics for evaluation datasets. This contract role is worldwide, requires English, and calls for 20 or more hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 16, 2026

Dutch-Speaking Personalization AI Response Evaluator

Evaluate personalized AI conversations in Dutch for helpfulness, accuracy, natural language, and evidence. This remote one-month contract pays $20 per hour and requires at least four hours per day.

Generative AI & RLHF
Text
Remote · Worldwide
Dutch
Part-time · Flexible
Entry level
Hourly · $20/hr

Posted Jul 16, 2026

Vietnamese-Speaking AI Personalization Evaluator

Review how conversational AI uses personal context in Vietnamese, compare responses, and explain your ratings. This remote contractor role pays $15 an hour for 20+ hours per week over three months.

Generative AI & RLHF
Text
Remote · Worldwide
Vietnamese
Part-time · Flexible
Intermediate level
Hourly · $15/hr

Posted Jul 16, 2026

Thai-Speaking AI Personalization Quality Analyst

Evaluate how well an AI uses personal context to make responses more helpful in Thai. This remote contract role pays $15 per hour for 20+ hours weekly and requires use of your primary personal account.

Generative AI & RLHF
Text
Remote · Worldwide
Thai
Part-time · Flexible
Intermediate level
Hourly · $15/hr

Posted Jul 16, 2026

Hindi-Speaking AI Personalization Quality Analyst

Evaluate Hindi AI conversations by writing prompts, comparing responses, checking personalization quality, and explaining your ratings. This contractor role pays $15 per hour and requires 20+ hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
Hindi
Part-time · Flexible
Entry level
Hourly · $15/hr

Posted Jul 16, 2026

Indonesian-Speaking AI Response Evaluator

Evaluate how well AI uses personal context in Indonesian conversations. Create multi-turn prompts, compare responses, and give clear feedback at $15 per hour with 20+ hours available each week.

Generative AI & RLHF
Text
Remote · Worldwide
Indonesian
Part-time · Flexible
Entry level
Hourly · $15/hr

Posted Jul 16, 2026

Korean-Speaking AI Personalization Response Evaluator

Create multi-turn Korean prompts and review AI responses for natural, helpful use of personal context. This remote contractor role pays $15 per hour and requires 20 or more hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
Korean
Part-time · Flexible
Intermediate level
Hourly · $15/hr

Posted Jul 16, 2026

Chinese-Speaking Personalized AI Response Evaluator

Evaluate how well conversational AI uses personal context in Chinese, compare responses, and write evidence-based feedback. This remote, three-month contract pays $15 per hour and requires at least 20 hours weekly.

Generative AI & RLHF
Text
Remote · Worldwide
Chinese
Part-time · Flexible
Entry level
Hourly · $15/hr

Posted Jul 16, 2026

Spanish-Speaking Personalized AI Response Evaluator

Evaluate Spanish conversational AI responses, design multi-turn prompts, and explain which answers are more helpful and accurate. This remote contractor role pays $15 per hour and requires 20+ hours weekly.

Generative AI & RLHF
Text
Remote · Worldwide
Spanish
Part-time · Flexible
Intermediate level
Hourly · $15/hr

Posted Jul 16, 2026

Arabic-Speaking AI Response Evaluator

Evaluate how well AI uses personal context in Arabic conversations. This remote three-month contract pays $15 per hour and requires at least four hours daily, with Pacific Time overlap.

Generative AI & RLHF
Text
Remote · Worldwide
Arabic
Part-time · Flexible
Intermediate level
Hourly · $15/hr

Posted Jul 16, 2026

Portuguese-Speaking AI Personalization Quality Analyst

Review how conversational AI uses personal context in Portuguese, create realistic multi-turn prompts, rank responses, and explain quality issues. This part-time contractor role pays $15 per hour and requires at least 20 hours each week.

Generative AI & RLHF
Text
Remote · Worldwide
Portuguese
Part-time · Flexible
Intermediate level
Hourly · $15/hr

Posted Jul 16, 2026

ML Evaluation Data Analyst

Analyze machine learning data, model outputs, evaluation metrics, and benchmark failures using Python and SQL. This remote contractor role requires 20 hours per week and is open in 10 countries.

Generative AI & RLHF
Text
Remote · India, Pakistan, Nigeria +7 more
English
Part-time · Flexible
Intermediate level

Posted Jul 16, 2026

What this work can involve

AI training and data labeling work turns human judgment into examples that help artificial-intelligence systems improve. For applicants located in Mexico, an individual role may focus on one defined type of task, such as checking whether content matches a guideline, transcribing or recording audio, reviewing model responses, translating text, or applying subject-matter knowledge. The live posting is the source of truth for the work involved and its requirements.

  • Possible tasks include image or video annotation, audio transcription, response rating, search-result evaluation, translation, and code review.
  • The specific project scope, qualifications, schedule, pay, and location rules are set by each individual posting.
  • Read the full posting before deciding whether a role matches your circumstances and eligibility.

How hiring works through OpenTrain

OpenTrain is the hiring organization for the public roles posted on OpenTrain. Candidates apply through OpenTrain, and OpenTrain may place or refer successful candidates to the client whose project they will support. Since requirements can differ between opportunities open to applicants in Mexico, use each live posting to confirm the applicable conditions.

  • Apply through OpenTrain rather than directly to an unnamed third-party platform.
  • Check the individual posting for its stated eligibility, qualifications, responsibilities, hours, pay, and location rules.
  • A role's acceptance of applicants in Mexico is determined by its own posting and may change over time.

Frequently asked questions

What can this work involve?
Remote AI training and data labeling work open to applicants located in Mexico can involve preparing or reviewing examples for artificial-intelligence systems. Depending on the posting, tasks may include annotating images or video, transcribing or recording audio, writing or rating model responses, evaluating search results, translating text, reviewing code, or applying subject-matter knowledge. Each live posting explains the actual responsibilities.
Where can I find role-specific details?
Find role-specific details in the individual live posting. That posting is the only source of truth for eligibility, responsibilities, qualifications, experience, pay, hours, schedule, location rules, and client-specific requirements. Do not assume that conditions listed for one opportunity apply to another role open to applicants located in Mexico.
How does hiring work through OpenTrain?
OpenTrain is the hiring organization for roles posted on OpenTrain, and candidates apply through OpenTrain. OpenTrain may place or refer successful candidates to the client whose project they will support. Review each live posting for the requirements and details that govern that particular opportunity.