Skip to content
OpenTrain AIFor AI Companies

STEM AI Agent Research Specialist

Review real AI-assisted STEM research sessions and evaluate reasoning, technical depth, and outputs. Remote contractor work pays $80-$100 per hour and requires 20+ hours per week.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $80–$100/hr

$80–$100/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Aug 4, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role and the #1 platform for finding and building careers in AI training and data labeling. OpenTrain helps contributors develop credible AI-training experience, discover projects that match their skills, and build a lasting professional portfolio.

  • Remote contractor engagement
  • Worldwide opportunity
  • English-language work

About AI Training Work

AI training is the human side of building modern artificial intelligence. Specialists review examples, assess model behavior, and provide expert feedback that helps AI systems become more accurate, useful, and reliable.

  • Work directly with cutting-edge AI systems
  • Use your technical expertise to shape model performance
  • Build experience in a fast-growing technology field

The Role

OpenTrain AI is seeking a STEM AI Agent Research Specialist to review historical session traces from local or agentic AI tools used for scientific research, engineering, coding, analysis, experimentation, and other technical problem-solving. You will examine how people guide these tools through complex work, validate the quality of their reasoning and outputs, and identify opportunities to improve AI performance.

This is a part-time contractor role for contributors available to work 20 or more hours per week. The engagement pays $80-$100 per hour.

  • Experience level: Entry level
  • Work type: Contractor and part-time
  • Pay: $80-$100 per hour
  • Minimum availability: 20+ hours per week

What You'll Do

You will assess technical AI sessions for authenticity, depth, reasoning quality, and usefulness as training material. The work involves documenting how research and problem-solving unfolded, identifying human intervention points, and communicating clear findings about AI behavior.

  • Review, curate, and submit historical session traces showing substantive STEM research or technical workflows
  • Determine whether sessions demonstrate multi-step reasoning, authentic human guidance, and sufficient technical depth
  • Document research methods, problem-solving approaches, human intervention points, outputs, and validation steps
  • Evaluate the effectiveness and limitations of agentic AI tools in STEM contexts
  • Highlight opportunities relevant to improving model performance
  • Write concise technical explanations and constructive feedback on AI accuracy, performance, and reasoning

Requirements

This role requires an advanced academic or professional background in science, technology, engineering, mathematics, or a related STEM discipline. You must also have hands-on experience using local or agentic AI tools for substantive research or technical projects beyond ordinary web-based chatbot use, as well as access to historical AI sessions involving substantial STEM or technical workflows.

  • Advanced academic or professional background in a STEM discipline
  • Hands-on experience guiding and validating multi-step outputs from local or agentic AI tools
  • Access to historical AI sessions involving substantive STEM or technical workflows
  • Strong scientific reasoning and technical research skills
  • Experience with data analysis, code debugging, or technical problem-solving
  • Strong research documentation and written communication skills
  • Ability to direct, challenge, test, and validate AI-generated outputs
  • Ability to communicate technical findings clearly in writing and conversation

Helpful Background

Experience in AI evaluation, output validation, iterative research methods, coding, experimentation, mathematics, physics, chemistry, biology, or life sciences can be valuable. Familiarity with local or agentic AI tools used for technical work is relevant.

  • AI evaluation or output validation
  • Iterative research methods
  • Coding and experimentation
  • Mathematics, physics, chemistry, biology, or life sciences
  • Claude Code, Claude Cowork, Codex, Claude for Life Sciences, or OpenCode

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all jobs

AI Safety Red Team Specialist

Probe conversational AI for jailbreaks, prompt injections, and misuse as a remote AI Safety Red Team Specialist for OpenTrain. Contractor, part-time role (20+ hrs/week) paying $24–$35/hr; native English and Thai required.

Generative AI & RLHF
Text
Remote · Worldwide
English, Thai
Part-time · Flexible
Intermediate level
Hourly · $24–$35/hr

Posted Jul 30, 2026

AI Safety Red Teaming Expert (English & Dutch)

Join OpenTrain AI as an expert red teamer probing conversational models for jailbreaks, prompt injection, bias exploitation, and multi-turn manipulation; remote contractor role, 20+ hrs/week, $48–$62/hr, native English and Dutch required.

Generative AI & RLHF
Text
Remote · Worldwide
English, Dutch
Part-time · Flexible
Expert level
Hourly · $48–$62/hr

Posted Jul 30, 2026

AI Safety Red Teamer

Probe conversational AI models and agents for jailbreaks, prompt injections, bias, and other safety failures. This expert, remote contract role offers flexible part-time work at $48–$62 per hour for fluent English and Swedish speakers.

Generative AI & RLHF
Text
Remote · Worldwide
English, Swedish
Part-time · Flexible
Expert level
Hourly · $48–$62/hr

Posted Jul 31, 2026