Build reinforcement learning environments that test AI models on real software engineering tasks. Work remotely on a flexible contract of fewer than 20 hours per week, earning $100-$150 per hour.
Coding & Software
100% Remote Hourly · $100–$150/hr
$100–$150/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Aug 5, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the #1 platform for finding and building careers in AI training and data labeling. We help people discover opportunities, build a profile, and grow their experience in a rapidly expanding field where human expertise directly shapes how AI systems learn and perform.
About AI Training Work
AI training is the human side of building artificial intelligence. Engineers, annotators, and evaluators create examples, tests, and feedback that help modern models reason more effectively, use tools correctly, and produce reliable results.
Work remotely with a computer and internet connection
Contribute directly to the development of cutting-edge AI systems
Choose flexible, part-time work that fits around other commitments
The Role
OpenTrain AI is recruiting a Software Engineer to create reinforcement learning environments for next-generation AI model training. You will design reproducible environments that test an AI model's ability to solve complex software engineering tasks using Model Context Protocol tools.
Your work will help measure both tool-use proficiency and engineering skill through deterministic verification and golden reference solutions. The role is listed as entry level in the project details, while the responsibilities and required capabilities call for substantial software engineering expertise.
Contractor position
Part-time schedule of fewer than 20 hours per week
Worldwide opportunity
English-language work
Pay: $100-$150 per hour
What You'll Do
You will build high-quality environments that give AI models realistic engineering challenges and provide accurate, repeatable ways to evaluate their solutions.
Design and implement Model Context Protocol tool interactions
Create tasks that require agents to discover and reason over information from real MCP servers
Develop deterministic verification methods and golden reference solutions
Ensure environments are reproducible and evaluate both MCP tool proficiency and software engineering ability
Collaborate with a remote team while maintaining high-quality output standards
Required Qualifications
This role is suited to a software engineer who can work across complex codebases, diagnose difficult problems, and create maintainable, high-performance solutions.
Proficiency in C++, Python, Java, GoLang, TypeScript, or Rust
Deep understanding of algorithms, data structures, and performance tuning
Demonstrated experience debugging complex software issues and delivering maintainable solutions
Strong background in feature development and codebase refactoring
Proven ability to optimize software for performance and scalability
Exceptional written and verbal communication skills with close attention to detail
Track record of success in collaborative, cross-functional teams, ideally in remote settings
Helpful Background
The following experience is helpful for creating realistic, technically rigorous training environments, though it is not listed as required.
Experience with large-scale, distributed codebases
Familiarity with modern AI or machine learning systems
Background in code reviews and software engineering best practices
Why This Work Matters
Reinforcement learning environments give AI systems structured ways to practice solving difficult problems and receive meaningful evaluation. By designing reliable coding tasks and verification systems, you will influence how future models reason about software engineering and use development tools.
Getting Started With OpenTrain
Creating an OpenTrain account is free. Build your profile, review opportunities that match your expertise, and apply in minutes for flexible AI training work that can grow into a lasting career.
Help evaluate and improve AI systems by building reinforcement learning environments for complex software engineering tasks. Work remotely for approximately 15 hours per week, with listed compensation of $50-$100 per hour.
Join OpenTrain as a part-time contractor to design reinforcement-learning code environments and expert reference solutions that teach AI to code, debug, refactor, and optimize — paid $50–$150/hr. Remote role with flexible hours; apply with a strong GitHub/GitLab record.
Build reproducible reinforcement learning environments that test how AI systems solve real-world coding problems. This remote contractor role offers flexible work of approximately 15 hours per week at $50–$100 per hour.