Use advanced medical expertise to write and verify challenging AI benchmark questions, solutions, and evidence-based references. This fully remote, asynchronous contract pays $94-$119 per hour.
Medical & Health
100% Remote Hourly · $94–$119/hr
$94–$119/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Aug 19, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. It helps contributors discover specialized projects, build a professional profile, and apply in minutes while developing a lasting portfolio of AI training work.
As an OpenTrain contractor, you will contribute your medical expertise to the human side of artificial intelligence. Creating and evaluating high-quality training material helps make advanced AI systems more accurate, reliable, and useful.
Free OpenTrain account and profile
Remote work with a global contributor community
Opportunities to build credible experience in AI training
About AI Training Work
AI models learn from carefully prepared examples and expert evaluations. In this project, your assessment questions, explanations, and reviews will help produce rigorous benchmarks for measuring what AI systems understand about medicine and health.
Apply specialist knowledge to cutting-edge AI evaluation
Help distinguish genuine clinical reasoning from superficial recall
Work remotely using a flexible, asynchronous model
The Role
OpenTrain is seeking a Medical AI Benchmark Content Expert to create and verify rigorous academic assessment material for an AI research initiative. The work spans clinical medicine and surgery, medical imaging and diagnostics, pharmacovigilance, healthcare management and economics, rehabilitation, and allied health.
Your subject-matter judgment will support gold-standard benchmarks for evaluating AI capabilities. The role is listed as entry level, but it requires advanced knowledge of medicine, biomedical sciences, public health, or a closely related health discipline.
Fully remote and asynchronous contract work
Compensation of $94-$119 per hour
Expected commitment of 20 or more hours per week
The role description specifies a minimum commitment of 10 or more hours per week
English-language work available worldwide
What You'll Do
You will author original, challenging multiple-choice questions that test deep conceptual understanding rather than surface recall. Each question must be self-contained and unambiguous, with all information needed to solve it included in the problem statement.
You will also review existing questions for accuracy, clarity, completeness, precision, and solvability. When changes are needed, you will make and document justified edits, assess the question's difficulty, and provide a complete evidence-based solution.
Write original medical assessment questions at Medium, Hard, or Expert difficulty
Provide one correct answer and nine plausible alternatives for each question
Explain the reasoning with clear, step-by-step solutions
Support each question with one to five reputable references
Use peer-reviewed journals or clinical guidelines as supporting sources
Evaluate question accuracy, clarity, completeness, precision, and solvability
Requirements
You should have advanced medical or health-science knowledge suitable for authoring rigorous assessment questions. Strong clinical reasoning, biomedical research methodology, and the ability to explain assessment decisions are essential.
Precise, self-contained question writing and excellent written English are required. You must be able to distinguish defensible answers from subtle distractors and judge whether a question can be solved from the information provided.
Advanced knowledge of medicine, biomedical sciences, public health, or a related health discipline
Clinical reasoning across a relevant medicine, biomedical science, public health, or allied health domain
Ability to identify correct answers among plausible distractors
Ability to assess question solvability and explain the reasoning behind decisions
Knowledge of biomedical research methodology
Excellent written English for precise academic communication
Helpful Background
An MD, DO, PhD, or doctoral candidacy in Medicine, Biomedical Sciences, Public Health, or a related field is well suited to this work. A master's degree may also be appropriate when paired with exceptional depth in a relevant subdomain.
Board certification
Clinical experience
Health-related research publications
Deep expertise in a relevant medical or allied health subdomain
Why Work With OpenTrain
OpenTrain gives freelancers one place to manage AI training opportunities and build a portfolio they control. A stronger profile can help you show credible experience, find projects aligned with your expertise, and grow AI training work into a long-term career.
Contribute directly to how state-of-the-art AI systems are evaluated
Turn specialized medical expertise into valuable AI training experience
Build a durable professional portfolio on OpenTrain
Review and improve AI-generated clinical responses as a medical specialist, focusing on diagnostic reasoning, patient safety, and treatment quality. This expert contract pays $150 per hour for 20+ hours weekly.
Use your senior clinical expertise to evaluate medical AI outputs, write gold-standard solutions, and design benchmarks that test clinical reasoning. This Bay Area engagement pays $70 to $110 per hour.
Help improve AI-generated medical and public health responses through clinical reasoning, fact-checking, written feedback, and response ranking. This remote contractor role offers part-time work under 20 hours per week and pays $25 to $72 per hour.