Lead large technical teams to deliver high-quality SFT and RLHF datasets for foundational language models. This part-time contractor role requires senior engineering leadership, hands-on quality judgment, and 20+ hours/week from candidates based in India, Brazil, Mexico, Argentina, Chile, Colombia,
Generative AI & RLHF
Remote
7 countries
Eligibility
Expert
Experience
Jul 20, 2026
Posted
Open to applicants in
India Brazil Mexico Argentina Chile Colombia Peru
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help people start and grow durable freelance careers teaching AI by hosting projects, tracking proof-of-work, and letting contributors build a unified profile that showcases their experience.
Working with OpenTrain means contracting directly on training projects, growing a credible portfolio, and accessing increasingly sophisticated LLM training work as you progress.
About AI training work
AI training (also called data labeling or human feedback work) is the human backbone of modern machine learning: people prepare, review, and refine examples that teach models how to behave. Projects range from annotating images and transcribing audio to writing, reviewing, and ranking model responses for supervised fine-tuning (SFT) and reinforcement learning from human feedback (RLHF).
This industry offers fully remote, flexible, and accessible roles—ideal for experienced engineers and managers who want to shape how state-of-the-art language models perform.
The role
OpenTrain is hiring an LLM Training Delivery Leader to own delivery quality, throughput, cost, and the transition of SFT and RLHF projects from launch to stable operations. You will lead large technical teams, partner with researcher clients and internal stakeholders, and take responsibility for operational excellence across technical and annotation pipelines.
This is a contract, part-time leadership role expected to run at 20+ hours per week and is open to applicants based in India, Brazil, Mexico, Argentina, Chile, Colombia, and Peru. English language fluency is required.
What you'll do
Lead and manage large teams of software engineers and data scientists delivering SFT and RLHF datasets.
Oversee supervised fine-tuning and reinforcement learning-from-human-feedback delivery pipelines and milestones.
Define and maintain rigorous review processes to protect dataset quality and ensure reproducible results.
Collaborate closely with researcher clients and internal stakeholders to align delivery to expectations and priorities.
Resolve technical and process issues related to code quality, annotation methodology, task clarity, and result communication.
Own the transition of projects from initiation to stable delivery operations, including throughput and cost responsibilities.
Requirements
8+ years of professional software engineering experience.
3+ years in an engineering management or delivery leadership role.
Hands-on technical judgment and experience identifying and resolving code quality and delivery issues.
Experience leading large teams in a delivery-oriented environment.
Proficiency with Python, Java, or JavaScript.
Strong communication and stakeholder management skills, including experience working with researcher clients.
Familiarity with SFT and RLHF workflows and the unique quality controls they require.
Available for at least 20 hours per week; hired as a contractor (part-time).
English fluency required; candidates must be based in India, Brazil, Mexico, Argentina, Chile, Colombia, or Peru.
Who should apply
Apply if you are an experienced engineering leader who has run delivery teams and cares deeply about dataset quality for language model training. This role suits managers who combine hands-on technical judgment with strong people and stakeholder skills.
If you have direct experience with supervised fine-tuning, RLHF, or managing annotation and model-training pipelines, you'll be especially competitive.
How it works
You will contract with OpenTrain AI on a part-time basis. Build or update your OpenTrain profile to show relevant engineering and LLM training experience—strong profiles help you demonstrate proof-of-work and win leadership engagements.
If selected, you will work remotely with distributed teams and researcher partners. OpenTrain handles project coordination; you focus on delivery leadership, quality, and operational outcomes.
Write realistic multi-turn conversations that teach assistants when and how to call tools, modeling calendars, email, maps, and other app workflows. Part-time contractor role requiring 20+ hrs/week and 3+ years technical or analytical experience; open to applicants in specified countries.
Join OpenTrain as an Insurance LLM Evaluation SME to design and score underwriting, claims, and risk-assessment evaluation tasks for LLMs. Remote (U.S. only), $60–$80/hr, 35 hours/week, contractor/part-time.
Join OpenTrain to design and solve advanced physics problems that probe LLM reasoning and symbolic skills; remote, part-time contractor work (20+ hrs/week) for candidates with graduate-level physics experience and strong written English.