Join OpenTrain AI to evaluate and improve AI-driven infrastructure automation: generate prompts, rate DevOps runbooks, and assess self-hosted platforms. Part-time contractor role requiring 2+ years DevOps experience, annotation/Q A skills, English B2+, 20+ hrs/week, $15–$45/hr.
Generative AI & RLHF
100% Remote Hourly · $15–$45/hr
$15–$45/hr
Compensation
Worldwide
Eligibility
Intermediate
Experience
Mar 29, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI hires and manages projects where people build careers teaching and improving artificial intelligence. We connect contributors with hands-on work that shapes how modern AI behaves while providing flexible, remote opportunities.
This role is offered directly by OpenTrain AI and is part of our program for training and evaluating models used in technical, infrastructure-focused scenarios.
About AI Training Work
AI training (also called data labeling, annotation, or human feedback) is the human side of building AI systems: people create, review, and rate examples so models learn useful, reliable behavior.
For technical and DevOps-focused projects, contributors write prompts, evaluate model troubleshooting guidance and runbooks, and rate outputs for correctness, safety, and operational reliability.
The Role
You will help train and evaluate AI systems that produce deployment and scaling strategies, operational runbooks, and troubleshooting steps for real-world infrastructure environments.
Work includes generating prompts, defining evaluation rubrics for uptime/monitoring/fault tolerance, reviewing and improving AI-generated DevOps runbooks, and assessing reliability and performance of self-hosted automation platforms under load.
What You'll Do
Create clear, scenario-based prompts for AI systems about deployment, scaling, and automated remediation.
Define and apply evaluation criteria and rubrics focused on uptime, monitoring, observability, and fault tolerance.
Review AI-written runbooks and troubleshooting steps for accuracy, completeness, and operational safety.
Assess self-hosted automation platforms for reliability and performance under load and document findings.
Contribute content and feedback for AI tutoring systems that teach operators how to manage complex infrastructure.
Provide structured evaluation ratings (EVALUATION_RATING) and written QA notes for text outputs.
Requirements
2+ years of hands-on DevOps, infrastructure, or backend systems experience.
Proven skill deploying and operating self-hosted environments and automation platforms.
Experience with monitoring, uptime, fault tolerance, and performance assessment under load.
Hands-on text annotation, evaluation, or rubric-based QA experience (you will rate and comment on text outputs).
Experience evaluating LLM outputs for legal reasoning quality (required).
English proficiency B2 or higher; CV must be in English and state your English level, email address, and phone number.
Who Should Apply
You are an intermediate-level engineer who enjoys translating infrastructure expertise into clear evaluation rubrics and actionable feedback for AI.
You have both operational experience (deployments, monitoring, fault tolerance) and prior annotation or QA work where you produced structured ratings and written critique.
Compensation, Time, and Logistics
This is a part-time contractor role requiring 20+ hours per week. Employment types: CONTRACTOR, PART_TIME.
Pay is hourly in USD with a range of $15–$45/hour (listed max $45/hr). Work involves text-based evaluation and rating (EVALUATION_RATING).
Data type: TEXT. Labeling: evaluation ratings and written QA.
Experience level: Intermediate.
You must submit a CV in English that includes your English proficiency level, an email address, and a phone number.
Eligibility & Restricted Locations
This role is open worldwide except for the locations listed below. Applicants must confirm eligibility when applying.
Restricted U.S. states: Alaska, Arkansas, California, Connecticut, Delaware, Georgia, Hawaii, Illinois, Indiana, Kansas, Louisiana, Maine, Maryland, Massachusetts, Nebraska, Nevada, New Hampshire, New Jersey, New Mexico, Ohio, Oregon, Tennessee, Utah, Vermont, Washington, West Virginia
Also restricted: Antarctica, Aruba, Åland Islands, Saint Barthélemy, Bonaire/Sint Eustatius and Saba, Bouvet Island, Cocos (Keeling) Islands, Democratic Republic of the Congo, Cook Islands, Christmas Island, Western Sahara, Falkland Islands (Malvinas), French Guiana, Guadeloupe, South Georgia and th
How To Apply
Submit your CV in English and indicate your English proficiency level; include an email address and phone number. Applications are handled by OpenTrain AI as the contracting organization.
When applying, highlight relevant DevOps projects, examples of annotation or rubric-based QA work, and experience evaluating LLM outputs for legal reasoning if available.
Join OpenTrain to build and evaluate RL training environments that teach AI systems how to operate cloud infrastructure. Remote contractor role, worldwide, 20+ hrs/week (see schedule), paying $60–$130/hr with estimated project earnings of $9,600–$20,800.
Lead quality assurance for AI-generated civil engineering content—review calculations, assumptions, units, and safety, give structured written feedback, and maintain QA materials. Contractor role (US only), remote, 20+ hrs/week, up to $105/hr.
Build and evaluate advanced n8n automations to train AI systems on reliable, scalable workflows; contractor role, 20+ hrs/week, paid $15–$45/hr (USD). Ideal for engineers with self‑hosted n8n, Docker, API integration, and rubric-based evaluation experience.