Create reproducible reinforcement learning environments that test how AI models solve real software engineering workflows. This worldwide remote contractor role offers approximately 15 hours per week and listed compensation of $100–$150 per hour.
Coding & Software
100% Remote Hourly · $100–$150/hr
$100–$150/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Aug 4, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover projects, build a professional profile, and grow their experience in a fast-moving field.
Creating an OpenTrain account is free. Your profile can help you present credible AI training experience, discover work aligned with your skills, and develop a long-term portfolio.
About AI Training and Model Evaluation
AI training is the human side of building artificial intelligence. People create examples, evaluate model behavior, and provide structured feedback that helps modern systems become more capable, reliable, and useful.
In this role, your software engineering expertise will help evaluate whether AI models can reason through realistic development, debugging, DevOps, and command-line workflows.
The Role
OpenTrain AI is recruiting a Software Engineering RL Environment Creator to design reproducible reinforcement learning environments for complex engineering tasks. You will turn realistic technical workflows into structured evaluations with clear inputs, expected behavior, and dependable golden reference solutions.
The work may involve Git, Docker, GDB, AddressSanitizer, FFmpeg, and related developer utilities. You will apply sound engineering judgment to make each environment technically meaningful, maintainable, and repeatable.
Worldwide remote contractor engagement
Approximately 15 hours per week
Part-time contract role
English-language work
Listed compensation: $100–$150 USD per hour; confirm compensation terms before accepting the engagement
What You’ll Do
You will create software engineering workflows that test an AI model’s ability to solve realistic technical problems. Your work will span development and evaluation, from designing the task environment through producing a reliable reference solution against which model-generated approaches can be assessed.
Design realistic DevOps, CI/CD, debugging, performance, and code-maintenance scenarios
Build reproducible environments with structured inputs and clearly defined expected behavior
Develop golden reference solutions for evaluating model-generated approaches
Use command-line and development tools where appropriate
Create workflows that are maintainable, repeatable, and technically meaningful
Apply engineering judgment to complex software and systems problems
Required Skills
This role is listed at the entry level, while requiring strong software engineering capability in the areas below. You must be proficient in at least one supported programming language and able to design, debug, and document complex engineering workflows.
Proficiency in at least one of C++, Python, Java, Go, TypeScript, or Rust
Strong understanding of algorithms, data structures, performance tuning, and scalable software design
Ability to design and debug complex DevOps, CI/CD, and software engineering workflows
Experience creating maintainable code and refactoring codebases
Ability to develop reliable golden or reference solutions
Demonstrated ability to debug complex software issues and deliver maintainable solutions
Careful written and verbal communication
Experience with collaborative software development
Helpful Background
The following experience is helpful but not required. It may be especially valuable when creating environments that represent complex production engineering challenges.
Large-scale or distributed codebases
Modern AI or machine learning systems
Rigorous code reviews
Software engineering best practices
Why Work With OpenTrain
AI training and data-labeling work is a rapidly growing part of the technology industry, and contributors directly influence how advanced AI systems behave. OpenTrain brings opportunities and professional growth tools together so you can build experience without starting from scratch for every project.
This flexible, remote engagement can fit around other commitments while giving you the opportunity to apply advanced software engineering skills to cutting-edge AI model evaluation.
Work remotely from anywhere in the world
Choose a part-time schedule of approximately 15 hours per week
Build experience at the intersection of software engineering and AI training
Create a stronger portfolio through your OpenTrain profile
Design and evaluate realistic coding tasks that train and test large language models for software engineering. Remote, contractor role (20+ hrs/week) for experienced engineers who can write tasks, tests, and review AI-generated code.
Design RL environments that test AI models on original competitive programming problems by building C++ checkers, reference solutions, and robust test suites. Remote, part-time contractor role (20+ hrs/week) paying $45–$65/hr for English speakers worldwide.
Evaluate how AI models solve complex software engineering problems by designing reproducible reinforcement learning environments, verification systems, and coding tasks. Remote, part-time contractor work pays $50–$100 per hour for less than 20 hours weekly.