Help improve medical AI by evaluating clinical reasoning, designing realistic scenarios, and assessing model decisions against evidence-based standards. This remote physician contract offers flexible scheduling for 20 to 30 hours per week.
Medical & Health
100% Remote
Worldwide
Eligibility
Entry
Experience
Aug 3, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. It helps people discover projects across the industry, build a professional profile, and apply in minutes. Creating an OpenTrain account is free.
Work directly with OpenTrain AI as a remote contractor.
Build experience in a fast-growing field at the intersection of medicine and artificial intelligence.
Create a profile that showcases your AI training and data-labeling work.
About AI Training Work
AI training is the human side of building artificial intelligence. Medical experts help evaluate whether AI systems produce sound, useful, and appropriately nuanced responses by applying the judgment and standards used in real clinical practice.
Contribute expert feedback that helps shape how advanced AI systems reason about medical problems.
Use your clinical knowledge in flexible, remote work that can fit around other commitments.
Help researchers identify where AI systems need stronger medical knowledge or reasoning.
The Role
OpenTrain AI is recruiting a Medical AI Clinical Reasoning Evaluator for a remote contract engagement. You will bring active clinical expertise to the assessment of medical reasoning and decision-making, helping research teams understand whether model outputs reflect the complexity and nuance required in clinical practice.
The engagement is initially one month, with potential extensions based on performance and fit. Scheduling is flexible, with availability of 20 or more hours per week and up to 30 hours per week. Compensation is not disclosed.
Role: Medical AI Clinical Reasoning Evaluator
Work arrangement: Remote contract engagement
Schedule: Flexible, 20 to 30 hours per week
Initial term: One month, with potential extensions
Language: English
Geography: Worldwide
What You'll Do
You will design systematic evaluation frameworks for medical AI systems and create realistic clinical scenarios that test reasoning and decision-making. You will also develop assessment methods that capture the nuance of clinical practice and support consistent evaluation.
Your work will include reviewing model behavior against realistic clinical problems and evidence-based standards, identifying gaps in medical knowledge and reasoning, and collaborating with AI researchers to improve model performance.
Design evaluation frameworks for medical AI systems.
Create realistic clinical scenarios that test AI reasoning and decision-making.
Develop assessment methods that reflect the nuance of clinical practice.
Evaluate model behavior against clinical problems and evidence-based standards.
Identify meaningful gaps in AI medical knowledge and reasoning.
Explain clinical judgments and assessment results clearly to research collaborators.
Requirements
You must be a licensed physician who is currently in active clinical practice. The role requires experience with clinical decision-making and evidence-based medicine, as well as the ability to evaluate complex medical reasoning and identify important gaps.
Active physician licensure and current clinical practice.
Experience applying clinical decision-making and evidence-based medicine.
Ability to assess clinical reasoning and identify gaps in medical knowledge.
Ability to design realistic clinical scenarios and evaluation frameworks.
Strong analytical skills for examining medical reasoning.
Clear communication skills for explaining complex clinical judgments and assessment results.
Interest in how AI can support clinical practice.
Helpful Clinical Background
Experience in the following areas can provide useful perspective when evaluating varied medical scenarios. Critical thinking and problem-solving skills are valuable for assessing model decisions across complex clinical contexts.
Public health
Pediatrics
Emergency medicine
Obstetrics and gynecology
Gynecological oncology
How to Apply Through OpenTrain
Create a free OpenTrain account, complete your profile, and apply in minutes. Your clinical expertise can help improve how AI systems approach medical problems while giving you experience in an emerging area of technology work.
Apply remotely through OpenTrain AI.
Highlight your physician licensure, active practice, and clinical decision-making experience.
Showcase relevant medical and AI training experience as your career develops.
Help evaluate how AI systems reason through real medical problems as a licensed physician. This flexible, remote one-month engagement offers up to 30 hours per week around your clinical commitments.
Use your senior clinical expertise to evaluate medical AI outputs, write gold-standard solutions, and design benchmarks that test clinical reasoning. This Bay Area engagement pays $70 to $110 per hour.
Use senior utilisation management and clinical review expertise to evaluate AI-generated medical-necessity decisions, care recommendations, and review summaries. This remote U.S. contract pays $100 to $150 per hour for 20+ hours weekly.