Skip to content
OpenTrain AIFor AI Companies

Backend Code Evaluation Engineer

Help improve AI systems by building backend coding environments, reference solutions, and rigorous evaluations. Remote contract work offers flexible scheduling, 20+ hours per week, and a target rate of $65-$120 per hour.

OpenTrain AI

Coding & Software

100% Remote Hourly · $65–$120/hr

$65–$120/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Aug 14, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the #1 platform for finding and building careers in AI training and data labeling. We help people discover projects, build a professional profile, and grow experience in a rapidly expanding field where human expertise shapes how AI systems work.

About AI Training Work

AI training is the human side of building artificial intelligence. Contributors write, review, and evaluate examples that help modern models become more accurate, capable, and reliable. This role applies that work to backend software development, combining engineering judgment with structured evaluation and technical feedback.

The Role

OpenTrain AI is hiring a Backend Code Evaluation Engineer for remote contract work focused on backend software problem solving and code evaluation. You will create realistic coding environments and reference solutions that help AI systems improve through structured human expertise, careful review, and clear technical feedback.

Prior AI training experience is not required. The estimated engagement is 20+ hours per week for approximately 1-3 months, with flexible scheduling.

  • Work arrangement: Remote, worldwide
  • Employment type: Contractor, part time
  • Target rate: $65-$120 per hour
  • Language: English
  • Experience level: Entry level

What You'll Do

You will design and contribute to coding tasks that test AI models on practical backend software problems. Your work will include creating reproducible environments, developing reference implementations, documenting technical reasoning, and reviewing peer submissions for correctness and clarity.

  • Create reinforcement learning environments for backend software problems.
  • Develop tasks involving code fixes, API development, database logic, data handling, and performance improvements.
  • Contribute code samples, debugging approaches, and practical development insights.
  • Build small to medium backend features.
  • Improve backend code and SQL or NoSQL database queries.
  • Create reproducible coding environments and golden reference solutions.
  • Document your reasoning and code decisions clearly.
  • Review peer-contributed code and technical submissions for correctness and clarity.

Requirements

You should have at least one year of hands-on experience with at least one of the listed programming languages and be comfortable analyzing unfamiliar backend codebases. Strong technical communication, documentation, and attention to detail are important for explaining decisions and making consistent review judgments.

  • At least one year of hands-on experience with Python 3, Java, Go, Rust, C++, or TypeScript.
  • Understanding of backend API design.
  • Understanding of SQL or NoSQL database fundamentals.
  • Ability to identify root causes and resolve backend bugs.
  • Ability to troubleshoot data-handling issues and performance bottlenecks.
  • Ability to create reproducible coding environments and golden reference solutions.
  • Ability to explain technical reasoning, code decisions, and review judgments clearly.

Why Work With OpenTrain

AI training and data labeling are among the fastest-growing ways to work in tech, with projects spanning software, language, images, audio, and model evaluation. This opportunity lets you apply backend engineering skills directly to the development of cutting-edge AI systems.

OpenTrain provides a place to build a lasting record of your AI training experience, discover relevant opportunities, and develop your work into a durable professional portfolio.

  • Work remotely from anywhere in the world.
  • Choose a flexible schedule around your availability.
  • Apply backend engineering expertise to state-of-the-art AI development.
  • Build experience in a growing technical field.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all jobs

AI Code Evaluation and Benchmarking Engineer

Evaluate and benchmark AI-generated code: review correctness, debug and verify solutions, and build evaluation datasets for frontier models. US-remote, contractor role — 20+ hrs/week (min 4 hrs/day), 1-month contract with 4-hour PST overlap required.

Coding & Software
Text
Remote · United States
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Software Engineering Code Evaluator

Evaluate and improve AI-generated code while creating benchmarking datasets for advanced software engineering models. This expert, remote contract role is part time at less than 20 hours per week.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Expert level

Posted Jul 16, 2026

Senior LLM Code Evaluation Engineer

Join OpenTrain AI to build evaluation datasets from public open-source code and measure how LLMs handle real-world software tasks; requires 3+ years software engineering with strong Go skills, 20+ hours/week, and eligibility in select countries.

Coding & Software
Computer Code Programming
Remote · India, Pakistan, Nigeria +6 more
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026