Skip to content
OpenTrain AIFor AI Companies

Code Quality Engineer for LLM Evaluation

Review AI-generated code for correctness, security, reliability, and maintainability in LLM evaluation work. This remote contract role requires at least seven years of software engineering experience and 20+ hours per week.

Apply now
OpenTrain AI

Coding & Software

100% Remote

Worldwide

Eligibility

Entry

Experience

Sep 30, 2026

Posted

Open worldwide

The Work

You will review AI-generated code used to train and evaluate large language models. Your work will cover different programming languages and software situations, including bug fixes, new features, refactoring, API integrations, configuration changes, and database operations.

You will assess whether code works as intended and identify defects, logic errors, missing pieces, edge-case failures, performance problems, security risks, and architecture weaknesses. You will compare possible implementations, improve code into reliable reference solutions, explain your recommendations, and help create technical rubrics and coding benchmarks.

  • Review AI-generated code and software changes for correctness, security, reliability, scalability, readability, and maintainability.
  • Debug complex codebases and identify the root causes of technical problems.
  • Compare implementations and select or create the solution that best meets the technical requirements.
  • Write clear technical feedback and contribute to evaluation criteria, coding benchmarks, and improved LLM evaluation methods.

What It Pays And Takes

The role details do not list a pay rate. This is a remote, part-time contract role for an individual contributor supporting AI training and engineering evaluation work.

  • Pay: Not provided in the role details.
  • Time: 20+ hours per week.
  • Location: Worldwide and remote.
  • Language: Written English proficiency is required.
  • Experience: The listing is marked entry level in the source fields, while the role description requires at least seven years of professional software engineering experience.
  • Programming: Strong proficiency in at least one of Python, JavaScript, TypeScript, Java, C++, Go, C#, Ruby, PHP, or Rust.
  • Technical knowledge: Production software development, debugging, code review, clean code, modular architecture, abstraction, error handling, data structures, algorithms, APIs, databases, and application architecture.
  • Practices: Experience with collaborative code reviews, Git, and modern software engineering methods.
  • Helpful background: Experience evaluating AI-generated code, creating technical rubrics, contributing to software engineering benchmarks, or working across multiple languages and architecture patterns.

How It Works

Apply on OpenTrain with your resume and then complete the application on the hiring site.

About AI Training Work

AI training is the human work behind systems that generate and understand code, text, images, and other data. People review examples, rate model outputs, and provide clear corrections so AI systems become more useful and reliable; OpenTrain helps people find and build careers in this field.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

LLM Code Evaluation Software Engineer

Review and improve AI-generated code, create coding datasets, and build tools that verify software solutions. This flexible contractor role is open to experienced engineers in the US, Canada, and selected Western European countries.

Coding & Software
Computer Code Programming
Remote · United States, Canada, Austria +3 more
English
Part-time · Flexible
Entry level

Posted Jul 16, 2026

LLM Code Evaluation Engineer

Review real open-source codebases, test bug-fixing tasks, and assess how large language models perform on software engineering problems. This remote contractor role requires 20+ hours per week and strong programming skills.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 16, 2026

C++ LLM Evaluation Software Engineer

Evaluate how well large language models solve real C++ software problems. Build code tasks from open-source projects, test repositories locally, and work remotely for 20 or more hours each week.

Coding & Software
Computer Code Programming
Remote · India, Pakistan, Nigeria +6 more
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026