Review real AI-assisted STEM research sessions and evaluate reasoning, technical depth, and outputs. Remote contractor work pays $80-$100 per hour and requires 20+ hours per week.
Generative AI & RLHF
100% Remote Hourly · $80–$100/hr
$80–$100/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Aug 4, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role and the #1 platform for finding and building careers in AI training and data labeling. OpenTrain helps contributors develop credible AI-training experience, discover projects that match their skills, and build a lasting professional portfolio.
Remote contractor engagement
Worldwide opportunity
English-language work
About AI Training Work
AI training is the human side of building modern artificial intelligence. Specialists review examples, assess model behavior, and provide expert feedback that helps AI systems become more accurate, useful, and reliable.
Work directly with cutting-edge AI systems
Use your technical expertise to shape model performance
Build experience in a fast-growing technology field
The Role
OpenTrain AI is seeking a STEM AI Agent Research Specialist to review historical session traces from local or agentic AI tools used for scientific research, engineering, coding, analysis, experimentation, and other technical problem-solving. You will examine how people guide these tools through complex work, validate the quality of their reasoning and outputs, and identify opportunities to improve AI performance.
This is a part-time contractor role for contributors available to work 20 or more hours per week. The engagement pays $80-$100 per hour.
Experience level: Entry level
Work type: Contractor and part-time
Pay: $80-$100 per hour
Minimum availability: 20+ hours per week
What You'll Do
You will assess technical AI sessions for authenticity, depth, reasoning quality, and usefulness as training material. The work involves documenting how research and problem-solving unfolded, identifying human intervention points, and communicating clear findings about AI behavior.
Review, curate, and submit historical session traces showing substantive STEM research or technical workflows
Determine whether sessions demonstrate multi-step reasoning, authentic human guidance, and sufficient technical depth
Document research methods, problem-solving approaches, human intervention points, outputs, and validation steps
Evaluate the effectiveness and limitations of agentic AI tools in STEM contexts
Highlight opportunities relevant to improving model performance
Write concise technical explanations and constructive feedback on AI accuracy, performance, and reasoning
Requirements
This role requires an advanced academic or professional background in science, technology, engineering, mathematics, or a related STEM discipline. You must also have hands-on experience using local or agentic AI tools for substantive research or technical projects beyond ordinary web-based chatbot use, as well as access to historical AI sessions involving substantial STEM or technical workflows.
Advanced academic or professional background in a STEM discipline
Hands-on experience guiding and validating multi-step outputs from local or agentic AI tools
Access to historical AI sessions involving substantive STEM or technical workflows
Strong scientific reasoning and technical research skills
Experience with data analysis, code debugging, or technical problem-solving
Strong research documentation and written communication skills
Ability to direct, challenge, test, and validate AI-generated outputs
Ability to communicate technical findings clearly in writing and conversation
Helpful Background
Experience in AI evaluation, output validation, iterative research methods, coding, experimentation, mathematics, physics, chemistry, biology, or life sciences can be valuable. Familiarity with local or agentic AI tools used for technical work is relevant.
AI evaluation or output validation
Iterative research methods
Coding and experimentation
Mathematics, physics, chemistry, biology, or life sciences
Claude Code, Claude Cowork, Codex, Claude for Life Sciences, or OpenCode
Probe conversational AI for jailbreaks, prompt injections, and misuse as a remote AI Safety Red Team Specialist for OpenTrain. Contractor, part-time role (20+ hrs/week) paying $24–$35/hr; native English and Thai required.
Join OpenTrain AI as an expert red teamer probing conversational models for jailbreaks, prompt injection, bias exploitation, and multi-turn manipulation; remote contractor role, 20+ hrs/week, $48–$62/hr, native English and Dutch required.
Probe conversational AI models and agents for jailbreaks, prompt injections, bias, and other safety failures. This expert, remote contract role offers flexible part-time work at $48–$62 per hour for fluent English and Swedish speakers.