Skip to content
OpenTrain AIFor AI Companies

RL Environment Developer

Join OpenTrain as an RL Environment Developer building reproducible reinforcement-learning environments that test AI models on complex software engineering workflows; contract, remote, 20+ hrs/week, $50–$150/hr based on experience.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $50–$150/hr

$50–$150/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Jul 20, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help people start and grow careers teaching AI by discovering projects, building a unified profile, and applying quickly — creating an OpenTrain account is free.

OpenTrain is the hiring and contracting organization for this role. We focus on real, practical AI-training work where contributors directly shape how state-of-the-art systems behave.

Why work in AI training

AI training (data labeling, annotation, RLHF and related work) is the human side of modern AI development. Contributors prepare and evaluate examples that models learn from, and this work is a flexible, accessible way to contribute to cutting-edge systems.

  • 100% remote: do this work from anywhere with a computer and internet connection.
  • Flexible, part-time contract work that can fit alongside other commitments.
  • Entry points exist for many backgrounds; specialized projects pay more for domain expertise.

The role — RL Environment Developer

We are recruiting an RL Environment Developer to design and implement reproducible reinforcement-learning environments that measure an AI model's ability to solve complex software engineering workflows. This is a contract, part-time role requiring 20+ hours per week and is open worldwide.

Compensation is hourly at $50–$150 per hour (USD), commensurate with experience. You will contribute production-quality code, golden reference solutions, reviews, and documentation to open-source repositories.

  • Employment type: Contractor, Part-time
  • Time commitment: 20+ hours/week
  • Pay: $50–$150 USD per hour, based on experience
  • Language: English required

What you'll do

  • Design and implement RL environments that simulate real-world software engineering workflows using CLI tools such as git, Docker, gdb, asan, ffmpeg, and others.
  • Create golden reference solutions for each environment to serve as ground truth for model evaluation.
  • Contribute production-quality code, reviews, and documentation to open-source repositories.
  • Optimize algorithms and system components using C++, Python, Java, GoLang, TypeScript, or Rust.
  • Identify and resolve technical challenges, bugs, and performance bottlenecks.
  • Collaborate with contributors and stakeholders to align deliverables with project goals.

Requirements

You must meet the core technical requirements below. We preserve all required qualifications from the role description.

  • Proven open-source contributions with a public GitHub or GitLab profile.
  • Proficiency in at least one of: C++, Python, Java, GoLang, TypeScript, or Rust.
  • Experience creating reproducible reinforcement-learning environments.
  • Familiarity with DevOps, CI/CD, and debugging tools and workflows (git, Docker, gdb, asan, ffmpeg).
  • Experience with large-scale distributed codebases and rigorous code reviews.
  • Strong problem-solving skills, attention to detail, and ability to design complex algorithms and system components.

Helpful background

  • Background in modern AI or machine learning systems.
  • Experience participating in code discussions and providing constructive feedback.
  • Previous work on software engineering evaluation, test generation, or similar evaluation tooling.

How it works / Apply

OpenTrain is the hiring organization for this role. To apply, create a free OpenTrain account, complete your profile, and submit your application with links to your public repositories and availability. Applications should highlight relevant open-source work and examples of reproducible environments or evaluation code.

  • Be prepared to share a public GitHub/GitLab profile and examples of past contributions.
  • Interviews or technical screenings may include code review or walkthroughs of environments you built.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar Jobs

View all jobs

AI Red Team Engineer, LLM Security

Join OpenTrain AI to red-team large language models and agents, building reproducible attack tests, automation, and security tooling. Part-time remote work for certified cybersecurity professionals with strong scripting and pentesting experience, flexible hours, pay up to $55/hr.

Generative AI & RLHF
Computer Code Programming
Remote · Worldwide
Part-time · Flexible
Intermediate level
Hourly · $40/hr

Posted Nov 17, 2025

LLM Red-Teamer for Adversarial Testing and Evaluation

Join OpenTrain as a remote LLM Red-Teamer to design adversarial multi-turn conversations, write rigorous evaluation rubrics, and validate frontier language models 20+ hrs/week for $40–$65/hr. Prior RLHF or evaluation experience is helpful but not required.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $40–$65/hr

Posted Jul 16, 2026

R Model Evaluation Engineer

Join OpenTrain as an R Model Evaluation Engineer to review AI-generated analyses, write expert R solutions, and improve model quality. Remote contract work at $55/hr, 20+ hours/week, for experienced R users who can judge statistical correctness and explain decisions clearly.

Generative AI & RLHF
Text
Remote · Bangladesh, Bhutan, Brazil +14 more
English
Part-time · Flexible
Intermediate level
Hourly · $55/hr

Posted Jul 9, 2026