Build and review software that helps improve AI models through coding evaluation, response ranking, fine-tuning datasets, and RLHF. This worldwide, part-time contract role offers 20+ hours per week for developers with Python and JavaScript or TypeScript expertise.
Coding & Software
100% Remote
Worldwide
Eligibility
Entry
Experience
Jul 17, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is recruiting contractors to contribute to cutting-edge model development, build a credible portfolio, and grow their experience in a fast-moving field. Creating an OpenTrain account is free.
Discover AI training work that matches your technical background
Build a profile that showcases your project experience
Apply in minutes and develop a long-term AI training portfolio
About AI Training Work
AI training is the human side of building artificial intelligence. Developers and evaluators help models improve by writing code, testing outputs, ranking responses, creating training examples, and applying consistent standards for accuracy, relevance, and safety.
This role combines software development with hands-on model evaluation. Your technical judgment will help shape how AI systems respond to coding and broader user-query tasks.
Work remotely with a computer and internet connection
Contribute to modern AI systems through practical technical feedback
Use flexible project work to build experience around your schedule
The Role
OpenTrain AI is seeking an AI Coding Evaluation and Training Developer for part-time contract work. You will design, develop, and maintain software used in AI model training and optimization while evaluating model performance and improving training data.
The role is worldwide, requires 20+ hours per week, and is classified as entry level. Strong technical skills and familiarity with AI evaluation workflows are central to the work.
Contractor position with part-time availability
Worldwide opportunity
20+ hours per week
English-language work
What You'll Do
You will combine coding, evaluation, documentation, and model-improvement tasks. The work requires clear reasoning, consistent application of defined criteria, and the ability to communicate technical judgments through constructive feedback and written rationales.
Design, develop, and maintain readable, reusable, efficient code for AI model training and optimization
Benchmark model performance and evaluate and rank responses against defined criteria
Document clear explanations and rationales for evaluation decisions
Create and maintain task-specific datasets for supervised fine-tuning
Collaborate on RLHF workflows and reward-model refinement
Develop evaluation strategies that support user needs, technical accuracy, and ethical alignment
Create and refine responses for model improvement
Review code and documentation, provide constructive feedback, and contribute to product enhancements
Required Skills And Background
You should be comfortable building production-quality software and assessing the quality of AI-generated responses. A bachelor's or master's degree in engineering or computer science, or equivalent experience, is preferred. Experience level is listed as entry level, but the role requires substantial technical proficiency across development and evaluation tasks.
Strong proficiency in Python and JavaScript or TypeScript, including ES6 conventions
Experience building modular, scalable web applications with Node.js or NestJS and React, Angular, or Vue.js
Ability to write clean, organized, correct, clearly annotated, readable, and maintainable code
Ability to evaluate and rank AI model responses using quality, relevance, and technical-accuracy criteria
Familiarity with supervised fine-tuning, RLHF, reward models, or AI model evaluation workflows
Docker knowledge and familiarity with software testing or quality assurance
Excellent spoken and written English communication
A bachelor's or master's degree in engineering or computer science, or equivalent experience, is preferred
Helpful Experience
The following background can help you contribute effectively, although it is listed as helpful rather than required.
Software quality assurance
Test planning or test-driven development
LLM prompting
Code review
Model evaluation
Supervised fine-tuning
Reinforcement learning with human feedback
Why Build An AI Training Career
Every major AI system depends on people who prepare examples, review outputs, and apply expert judgment. Coding evaluators are part of this human layer, helping models become more accurate, useful, and aligned with user needs.
OpenTrain helps you discover opportunities, present your experience in one profile, and turn project work into a durable AI training portfolio. Apply through OpenTrain to take the next step in this rapidly growing industry.
Work at the intersection of software engineering and artificial intelligence
Gain practical experience with evaluation, fine-tuning, and RLHF workflows
Build a profile that supports continued growth in AI training and data labeling
Evaluate AI-generated C and C++ code for correctness, efficiency, reliability, and maintainability while helping improve dialogue agents and LLM systems. This flexible contract role requires 20+ hours per week.
Review complete AI-assisted coding sessions for correctness, reasoning quality, workflows, and engineering practices. This remote U.S. contract role pays $70-$90 per hour and requires three years of software development experience.
Use your software engineering expertise to create coding challenges, reference solutions, and evaluations that improve AI systems. This remote contractor role offers flexible work of about 15 hours per week and listed pay of $50 to $100 per hour.