Skip to content
OpenTrain AIFor AI Companies

Biology Model Evaluation Expert

Join OpenTrain AI to design challenging biology problems and write rigorous step-by-step solutions that probe large language model reasoning; remote, contractor role for experts in biology (20+ hours/week, English required).

OpenTrain AI

Generative AI & RLHF

100% Remote

Worldwide

Eligibility

Expert

Experience

Jul 17, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the leading platform for building careers in AI training and data labeling. We connect expert contributors with meaningful evaluation and annotation work so they can build a durable freelance career and a unified AI training portfolio they control.

OpenTrain AI is the hiring and contracting organization for this role. We run projects that help shape how state-of-the-art AI systems are tested and improved, and we support contributors with clear project goals and collaboration tools.

About AI training and evaluation work

AI training (data labeling and human evaluation) is the human side of building intelligent systems: people create examples, judge outputs, and define benchmarks that teach models to reason and follow instructions. This work is remote, flexible, and directly influences how models behave.

As a biology model evaluator you will work at the intersection of life-sciences expertise and model evaluation, designing problems that reveal conceptual gaps and improve model safety and reliability.

The role

We are recruiting a Biology Model Evaluation Expert to design and solve advanced biology problems for large language model evaluation. This role focuses on creating problems and rigorous step-by-step solutions that test conceptual abstraction, multi-step reasoning, and data interpretation.

This is a remote contractor, part-time role requiring 20+ hours per week. Work language: English. Open to contributors worldwide.

  • Data type: TEXT; label types: EVALUATION_RATING and TEXT_GENERATION
  • Employment: CONTRACTOR, PART_TIME
  • Time requirement: 20+ hours/week
  • Language: English; worldwide applicants welcome

What you'll do

  • Design biology problems that probe large language model limitations across undergraduate to PhD-level topics
  • Write high-quality, rigorous step-by-step solutions that demonstrate correct reasoning and expected outputs
  • Collaborate with the evaluation team to align problem types with benchmark goals and identify failure modes
  • Help define and document new biology evaluation benchmarks spanning curricula and research-level topics
  • Provide constructive, detailed annotations and feedback on model outputs and on problem/solution quality

Requirements

  • Advanced biology knowledge spanning undergraduate through PhD-level topics (Biology, Biotechnology, Biochemistry, or related)
  • Proven ability to design difficult, multi-step problems that reveal reasoning gaps in LLMs
  • Skill writing clear, rigorous, step-by-step solutions and explanations
  • Strong English comprehension and structured written communication
  • Comfort giving detailed feedback and annotations in a remote collaboration setting
  • Master's, Ph.D., or postdoctoral study in a biology-related discipline preferred

Who should apply

This role is for subject-matter experts who enjoy turning technical mastery into clear test problems and explanations. Ideal applicants are researchers, instructors, or experienced practitioners comfortable with curriculum-level and research-level biology.

You should be detail-oriented, enjoy remote collaboration, and be motivated by improving model behavior and robustness through careful evaluation design.

How it works

After you join, you'll receive evaluation goals and templates. Typical tasks include drafting problem prompts, producing canonical solutions, annotating model responses, and iterating with the team on rubrics and benchmark scope.

Work is managed remotely via OpenTrain tools; you'll submit problem sets and solutions, complete annotation tasks, and participate in review cycles. Compensation details are set per project and shared during onboarding.

  • Create an OpenTrain account to apply and manage your work and portfolio
  • Work remotely and on a flexible schedule consistent with the 20+ hour weekly expectation
  • Contributions feed into public-facing benchmarks and internal evaluation suites that shape future model development

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar Jobs

View all jobs

Biology Expert for AI Training

Join OpenTrain to apply your biology expertise to train and evaluate AI models: write expert responses, review model outputs, and provide clear scientific feedback. Remote, part-time contractor work paying $80–$90/hr for roughly 20 hours/week over 1–3 months.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $80–$90/hr

Posted Jun 28, 2026

Biology LLM Evaluation Expert

Design and solve challenging biology problems to probe and evaluate large language models, creating step-by-step solutions and benchmark material. Remote contractor role (20+ hrs/week), worldwide, English required — OpenTrain AI hires directly.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Expert level

Posted Jul 17, 2026

Computational Biology AI Evaluation Expert

Use your computational biology expertise to evaluate and annotate AI-generated outputs across genomics, structural biology, and systems biology. Remote contractor role, 20+ hrs/week, $40–$60/hr — help shape safer, more accurate scientific AI.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Expert level
Hourly · $40–$60/hr

Posted Jul 7, 2026