Skip to content
OpenTrain AIFor AI Companies

AI Coding Agent Evaluation Expert

Evaluate AI coding agents using your professional STEM experience, document complex technical workflows, and provide feedback that improves next-generation coding systems. Remote contractor work pays $80–$100 per hour and requires 20+ hours weekly.

OpenTrain AI

Coding & Software

100% Remote Hourly · $80–$100/hr

$80–$100/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Aug 7, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the #1 platform for finding and building careers in AI training and data labeling. We help people discover projects, build a professional profile, and grow experience in a rapidly expanding field where human expertise directly shapes how AI systems work.

  • Remote contractor opportunity open worldwide
  • Build a portfolio of experience in AI training and evaluation
  • Create an OpenTrain account for free

About AI Training Work

AI training is the human side of building artificial intelligence. Experts review, evaluate, explain, and improve AI-generated work so that modern systems become more accurate, useful, and reliable. This project focuses on evaluating coding agents through realistic technical problem-solving and professional workflows.

  • Work directly on cutting-edge AI systems
  • Use your existing technical expertise rather than prior AI-training experience
  • Contribute remotely with flexible project-based work

The Role

OpenTrain AI is recruiting an AI Coding Agent Evaluation Expert for a remote contractor project. You will use hands-on experience with AI coding agents and professional STEM work to document technical workflows, assess AI-generated work, and provide structured feedback that improves coding agents.

This role is listed at the entry level, and previous AI-training experience is not required. However, you must have active professional experience in a STEM discipline and practical experience using AI coding agents beyond basic code generation or autocomplete.

  • Employment type: Contractor, part-time
  • Workload: 20+ hours per week
  • Location: Remote, worldwide
  • Working language: English

What You'll Do

You will capture detailed examples of how AI coding agents are used in demanding technical settings. Your work will help project teams understand where these systems perform well, where they struggle, and how their outputs can become more accurate and usable.

  • Document technical challenges involving research, engineering, coding, debugging, large-codebase navigation, or architectural decisions
  • Describe the tools, strategies, and decision-making processes used in technical workflows
  • Explain how AI coding assistance affected the work and its outcomes
  • Evaluate the practical effectiveness of AI agents in complex technical problem-solving
  • Provide candid feedback on AI coding agents' strengths and limitations
  • Review, test, and validate AI-generated technical work
  • Collaborate with project facilitators to improve technical accuracy and usability

Requirements

The strongest candidates combine professional STEM experience with meaningful, hands-on use of AI coding agents. You should be comfortable explaining technical decisions clearly, assessing AI-generated work carefully, and working through ambiguous problems independently.

  • Active professional experience in coding, data, engineering, scientific research, or another STEM discipline
  • Hands-on use of AI coding agents for multi-step STEM or technical work beyond basic code generation or autocomplete
  • Ability to explain technical concepts, workflows, and decisions clearly in writing and conversation
  • Sound judgment when reviewing, testing, and validating AI-generated technical work
  • Ability to navigate complex codebases or systems
  • Ability to solve ambiguous technical problems and provide constructive feedback on accuracy and usability
  • Ability to work independently, follow detailed guidelines, and collaborate effectively

Project Details and Compensation

This is a remote contractor opportunity paying $80–$100 per hour. You should be available for 20 or more hours per week. Project scheduling and duration will be confirmed for the engagement.

  • Pay: $80–$100 per hour
  • Schedule: 20+ hours per week
  • Engagement: Part-time contractor project
  • Project timing and duration: Confirmed for the engagement

Why Work With OpenTrain

OpenTrain gives you a place to build a credible record of AI-training and evaluation experience while discovering projects aligned with your skills. Your work on this project can help shape how AI coding agents support research, engineering, software development, and other complex technical work.

  • Apply your professional STEM expertise to real AI improvement work
  • Work remotely with a flexible part-time structure
  • Build a lasting profile in a fast-growing AI training industry
  • Find and manage opportunities through one professional OpenTrain profile

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar Jobs

View all jobs

AI Coding Agent Evaluation Expert

Evaluate AI coding agents on complex software development workflows, from debugging and codebase navigation to architectural changes. This remote contractor role pays $80–$100 per hour and requires 20+ hours per week.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $80–$100/hr

Posted Aug 7, 2026

AI Coding Agent Evaluation Expert

Help improve AI coding assistants by evaluating complex technical work, from large-codebase debugging to architectural changes and end-to-end features. This worldwide, part-time contractor role offers $80–$100 per hour and requires 20+ hours weekly.

Coding & Software
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $80–$100/hr

Posted Aug 7, 2026

Agentic Coding Annotator, Model Evaluation

Evaluate and improve agentic coding models by reviewing agent trajectories, verifying outputs, and designing rubrics on a 5-week remote contract. Requires 5+ years of hands-on software experience and daily overlap with PST.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 29, 2026