Skip to content
OpenTrain AIFor AI Companies

Biology and Health AI Evaluation Specialist

Apply now

Biology and Health AI Evaluation Specialist

Create and evaluate advanced text-only biology and health tasks that help improve AI models. This remote, four-week contractor engagement requires graduate-level expertise, at least 20 hours weekly, and Pacific Time overlap.

OpenTrain AI

Medical & Health

100% Remote

Worldwide

Eligibility

Entry

Experience

Sep 11, 2026

Posted

Open worldwide

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping contributors discover projects, build a professional profile, and apply in minutes.

This engagement gives you the opportunity to apply your biology and health expertise to advanced AI evaluation work while contributing to a portfolio in a fast-growing field.

About AI Training and Evaluation

AI training is the human side of building artificial intelligence. Specialists create examples, review model responses, and identify errors so AI systems can produce more accurate, useful, and reliable results.

In this role, your scientific judgment will help evaluate text-based questions and answers, diagnose model failure modes, and improve the quality of datasets used to train and assess AI systems.

The Role

OpenTrain is recruiting a Biology and Health AI Evaluation Specialist to create and review advanced assessment tasks for AI models. You will work with text-only biology problems and domain-specific question-and-answer tasks, applying rigorous scientific, academic, and clinical standards.

The work combines expert problem authoring with careful evaluation of model-generated content. You will refine tasks to meet demanding difficulty targets while maintaining scientific correctness, precise terminology, and clear explanations.

  • Remote contractor engagement
  • Four-week engagement
  • Part-time participation supported
  • At least 20 hours per week and up to 40 hours per week
  • At least 4 hours per day
  • Four hours of overlap with Pacific Time required
  • Weekend on-call availability required
  • Personal computer and stable, high-speed internet required

What You’ll Do

You will develop and curate challenging biology and health evaluation content for AI training datasets. The work requires translating complex scientific concepts into precise, self-contained text-only tasks and assessing whether AI-generated responses meet demanding standards.

  • Author upper-undergraduate and graduate-level biology problems.
  • Produce clear, fully worked solutions.
  • Convert diagram-heavy or multi-part concepts into self-contained text-only tasks.
  • Review biology and health questions and answers for scientific correctness.
  • Check terminology precision, clarity, and appropriate difficulty.
  • Diagnose AI-generated errors and identify model failure modes.
  • Refine questions and evaluation tasks for calibrated difficulty.
  • Collaborate with others to maintain consistent quality standards across large datasets.

Required Qualifications

A master’s degree or PhD in biology, health sciences, medicine, or a closely related biological field is required. The role also requires strong knowledge of biological systems, health concepts, and clinical or research-based reasoning.

You must be able to write precise scientific explanations in fluent English and judge complex questions and answers for correctness, terminology, clarity, and difficulty. Prior experience teaching, grading, creating university-level examinations, or performing research-based analytical work is required.

  • Graduate-level knowledge of biological systems and health concepts.
  • Clinical or research-based reasoning ability.
  • Experience authoring advanced biology problems with fully worked solutions.
  • Ability to translate complex concepts into precise text-only tasks.
  • Experience evaluating scientific questions and answers.
  • Ability to diagnose AI-generated errors and improve task quality.
  • Fluent written English and strong scientific explanation skills.
  • Familiarity with AI training data best practices is useful.
  • Prior AI evaluation experience can support the work.

Why Build an AI Training Career With OpenTrain

AI training and data-labeling work is a growing way to contribute to technology without leaving your area of expertise. Projects may involve evaluating model outputs, writing examples, reviewing specialized content, or improving the data behind modern AI systems.

OpenTrain helps you build a durable AI training portfolio by giving you one place to manage opportunities, demonstrate credible experience, and find work aligned with your skills. Creating an OpenTrain account is free.

  • Remote work with flexible part-time participation.
  • Directly apply scientific expertise to cutting-edge AI development.
  • Build experience in model evaluation and AI training.
  • Create a professional portfolio that reflects your specialized capabilities.

Engagement Details and Application

This is a remote contractor engagement focused on advanced biology and health evaluation work. The schedule requires at least 4 hours per day, with a weekly commitment of 20 or more hours and a maximum of 40 hours. Four hours of Pacific Time overlap and weekend on-call availability are required.

Apply through OpenTrain to be considered for this specialized AI evaluation opportunity.

  • Engagement type: Contractor and part-time
  • Duration: 4 weeks
  • Work format: Remote
  • Language: English
  • Data format: Text
  • Task types: Text generation, question answering, and evaluation rating

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Biostatistics AI Evaluation Expert

Use advanced biostatistical expertise to create rigorous AI evaluation tasks covering clinical trials, regulatory reviews, and observational studies. Contract work pays $60-$100 per hour for 20+ hours weekly.

Medical & Health
Document
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $60–$100/hr

Posted Aug 12, 2026

Medical Evaluation Specialist

Use clinical knowledge to create challenging medical questions, verify answers against literature and guidelines, and evaluate AI responses. This flexible worldwide contract offers 10-15 hours per week at $40-$90 per hour.

Medical & Health
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $40–$90/hr

Posted Aug 21, 2026

Clinical Trial Biostatistics AI Evaluator

Use clinical-trial biostatistics expertise to evaluate AI-generated analyses, verify statistical outputs, and create reliable clinical research evaluation tasks. This remote contract offers $60-$65 per hour with a 20+ hour weekly commitment.

Medical & Health
Document
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $60–$65/hr

Posted Aug 4, 2026