Skip to content
OpenTrain AIFor AI Companies

Backend Engineer, Cloud Infrastructure AI Evaluation

Build cloud infrastructure environments that test how well AI systems design, deploy, troubleshoot, secure, scale, and recover production systems. This contract role pays $50 to $100 per hour and requires 20+ hours weekly.

Apply now
OpenTrain AI

Coding & Software

Remote Hourly · $50–$100/hr

$50–$100/hr

Compensation

230 countries

Eligibility

Entry

Experience

Aug 5, 2026

Posted

Open to applicants in

United States
+ more
  • Åland Islands
  • Albania
  • Algeria
  • American Samoa
  • Andorra
  • Angola
  • Anguilla
  • Antarctica
  • Antigua & Barbuda
  • Argentina
  • Armenia
  • Aruba
  • Australia
  • Austria
  • Azerbaijan
  • Bahamas
  • Bahrain
  • Bangladesh
  • Barbados
  • Belgium
  • Belize
  • Benin
  • Bermuda
  • Bhutan
  • Bolivia
  • Bosnia & Herzegovina
  • Botswana
  • Bouvet Island
  • Brazil
  • British Indian Ocean Territory
  • British Virgin Islands
  • Brunei
  • Bulgaria
  • Burkina Faso
  • Burundi
  • Cambodia
  • Cameroon
  • Canada
  • Cape Verde
  • Caribbean Netherlands
  • Cayman Islands
  • Central African Republic
  • Chad
  • Chile
  • Christmas Island
  • Cocos (Keeling) Islands
  • Colombia
  • Comoros
  • Congo - Brazzaville
  • Cook Islands
  • Costa Rica
  • Côte d’Ivoire
  • Croatia
  • Curaçao
  • Cyprus
  • Czechia
  • Denmark
  • Djibouti
  • Dominica
  • Dominican Republic
  • Ecuador
  • Egypt
  • El Salvador
  • Equatorial Guinea
  • Eritrea
  • Estonia
  • Eswatini
  • Ethiopia
  • Falkland Islands (Islas Malvinas)
  • Faroe Islands
  • Fiji
  • Finland
  • France
  • French Guiana
  • French Polynesia
  • French Southern Territories
  • Gabon
  • Gambia
  • Georgia
  • Germany
  • Ghana
  • Gibraltar
  • Greece
  • Greenland
  • Grenada
  • Guadeloupe
  • Guam
  • Guatemala
  • Guernsey
  • Guinea
  • Guinea-Bissau
  • Guyana
  • Haiti
  • Heard & McDonald Islands
  • Honduras
  • Hungary
  • Iceland
  • India
  • Indonesia
  • Ireland
  • Isle of Man
  • Israel
  • Italy
  • Jamaica
  • Japan
  • Jersey
  • Jordan
  • Kazakhstan
  • Kenya
  • Kiribati
  • Kosovo
  • Kuwait
  • Kyrgyzstan
  • Laos
  • Latvia
  • Lebanon
  • Lesotho
  • Liberia
  • Liechtenstein
  • Lithuania
  • Luxembourg
  • Madagascar
  • Malawi
  • Malaysia
  • Maldives
  • Mali
  • Malta
  • Marshall Islands
  • Martinique
  • Mauritania
  • Mauritius
  • Mayotte
  • Mexico
  • Micronesia
  • Moldova
  • Monaco
  • Mongolia
  • Montenegro
  • Montserrat
  • Morocco
  • Mozambique
  • Namibia
  • Nauru
  • Nepal
  • Netherlands
  • New Caledonia
  • New Zealand
  • Nicaragua
  • Niger
  • Nigeria
  • Niue
  • Norfolk Island
  • North Macedonia
  • Northern Mariana Islands
  • Norway
  • Oman
  • Pakistan
  • Palau
  • Palestine
  • Panama
  • Papua New Guinea
  • Paraguay
  • Peru
  • Philippines
  • Pitcairn Islands
  • Poland
  • Portugal
  • Puerto Rico
  • Qatar
  • Réunion
  • Romania
  • Rwanda
  • Samoa
  • San Marino
  • São Tomé & Príncipe
  • Saudi Arabia
  • Senegal
  • Serbia
  • Seychelles
  • Sierra Leone
  • Singapore
  • Sint Maarten
  • Slovakia
  • Slovenia
  • Solomon Islands
  • South Africa
  • South Georgia & South Sandwich Islands
  • South Korea
  • Spain
  • Sri Lanka
  • St. Barthélemy
  • St. Helena
  • St. Kitts & Nevis
  • St. Lucia
  • St. Martin
  • St. Pierre & Miquelon
  • St. Vincent & Grenadines
  • Suriname
  • Svalbard & Jan Mayen
  • Sweden
  • Switzerland
  • Taiwan
  • Tajikistan
  • Tanzania
  • Thailand
  • Timor-Leste
  • Togo
  • Tokelau
  • Tonga
  • Trinidad & Tobago
  • Tunisia
  • Türkiye
  • Turkmenistan
  • Turks & Caicos Islands
  • Tuvalu
  • U.S. Outlying Islands
  • U.S. Virgin Islands
  • Uganda
  • United Arab Emirates
  • United Kingdom
  • United States
  • Uruguay
  • Uzbekistan
  • Vanuatu
  • Vatican City
  • Vietnam
  • Wallis & Futuna
  • Western Sahara
  • Zambia
  • Zimbabwe

The work

You will build reinforcement learning environments for training and evaluating AI models. These environments must reflect realistic cloud operations and produce reproducible results.

You will combine backend engineering, cloud architecture, DevOps automation, and technical evaluation to test AI performance in production-style scenarios.

  • Design environments that assess AI systems on infrastructure design, deployment, troubleshooting, security, scaling, and recovery.
  • Create scenarios involving distributed systems, networking, identity and access management, message queues, durable storage, observability, rolling deployments, and disaster recovery.
  • Develop deterministic validation tests and reference solutions for consistent assessment.
  • Build defective variants and failure scenarios that test how models respond and recover.
  • Document architecture, edge cases, and operational flows so environments are clear and reproducible.
  • Refine environment specifications and acceptance criteria with technical collaborators.
  • Use infrastructure automation and DevOps practices to deliver scalable, secure, maintainable evaluation systems.

What it pays and takes

This role is listed as entry level, but the work requires strong backend and cloud infrastructure skills. Prior AI experience is not required when you have the required backend and cloud expertise.

  • Pay: $50 to $100 USD per hour.
  • Time: 20+ hours per week.
  • Work type: Part-time contractor.
  • Language: English.
  • Location: Applicants must be in a country included in the listing; this role is not marked worldwide.
  • Programming: Strong experience with C++, Python, Rust, Go, Java, or JavaScript.
  • Infrastructure: Practical experience with DevOps, cloud infrastructure, CI/CD pipelines, and automation tools.
  • Systems: Ability to architect, scale, and secure distributed systems in production-grade environments.
  • Technical knowledge: Networking, IAM, queues, durable storage, observability, deployments, and disaster recovery.
  • Evaluation: Strong judgment when designing failure scenarios, acceptance criteria, and deterministic technical tests.

How it works

Apply on OpenTrain with your resume, then complete the application on the hiring site.

About AI training work

AI training work is the human work behind systems that learn from examples, including testing model behavior in realistic technical environments. People with strong specialist skills are needed to create reliable tests and judge whether an AI system handles complex tasks correctly.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

AI Infrastructure Automation Engineer

Help train and evaluate AI systems for infrastructure design, DevOps automation, reliability, and high-volume workflows. This worldwide contract role pays $15 to $45 per hour and requires 20+ hours weekly.

Coding & Software
Text
Remote · Worldwide
Part-time · Flexible
Intermediate level
Hourly · $15–$45/hr

Posted Mar 29, 2026

AI Data Evaluation QC Engineer

Review AI evaluation tasks, outputs, rubrics, verifiers, and test cases for technical quality. This remote, three-month contractor role requires at least 20 hours per week and four hours of Pacific Time overlap.

Coding & Software
Text
Remote · India, Pakistan, Nigeria +6 more
English
Part-time · Flexible
Entry level

Posted Sep 21, 2026

AI Evaluation Engineer, Engineering Simulation

Build and validate demanding engineering simulations that test AI agents across electrical, mechanical, control systems, aerospace, systems, and robotics applications. This remote contractor role requires strong Python, simulation, and engineering design experience.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Sep 11, 2026