Use your psychiatry expertise to challenge and improve large language models through advanced diagnostic, clinical reasoning, and risk-assessment evaluations. This remote contract role pays $150 per hour and is open to qualified professionals in the US, UK, and Australia.
Medical & Health
Remote Hourly · $150/hr
$150/hr
Compensation
3 countries
Eligibility
Entry
Experience
Aug 15, 2026
Posted
Open to applicants in
United States United Kingdom Australia
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the #1 platform for finding and building careers in AI training and data labeling. We help contributors discover specialized projects, build a professional profile, and apply to opportunities that shape how modern AI systems work.
About AI Training Work
AI training is the human side of building artificial intelligence. Expert contributors evaluate model responses, identify reasoning failures, and create high-quality examples that help AI systems become more accurate, reliable, and useful.
This role applies specialist psychiatric knowledge to advanced language-model evaluation. Your clinical judgment will help assess how models handle complex cases, nuanced differentials, and safety-sensitive decisions.
The Role
OpenTrain AI is recruiting a Psychiatric AI Evaluation Expert to evaluate how large language models reason through advanced psychiatric cases. You will design expert-level diagnostic and reasoning challenges that current models fail, then provide authoritative, gold-standard answers.
The work focuses on edge cases, nuanced differentials, and risk assessment. You will also collaborate directly with research teams at a leading AI lab and provide feedback on common failure modes in model reasoning.
Contractor position
Part-time engagement of 20+ hours per week
Pay: $150 USD per hour
Remote work available in the United States, United Kingdom, or Australia
English-language work
What You'll Do
Develop gold-standard answers with clear clinical reasoning.
Create advanced psychiatric diagnostic and reasoning challenges.
Emphasize edge cases, nuanced differentials, and risk assessment.
Evaluate large language model reasoning on complex psychiatric cases.
Collaborate directly with research teams at a leading AI lab.
Identify and explain common failure modes in model reasoning.
Required Qualifications
This is an expert-level evaluation role requiring completed psychiatric training, current professional standing, and substantial independent clinical experience. International medical degrees and national equivalents are accepted where specified.
Completed psychiatry residency through ACGME, AOA, or a national equivalent.
Medical degree: MD or DO; international degrees qualify.
Active board certification in psychiatry through ABPN, FRCPC, MRCPsych, FRANZCP, or an equivalent body.
Active, unrestricted medical license with no suspension or open disciplinary action.
At least three years of independent psychiatric practice after residency.
Current practice in an accredited or government-licensed setting.
Experience developing clinical reasoning challenges or serving as an OSCE examiner.
Based in the US, UK, or Australia.
Helpful Background
At least 10 direct patient-contact hours per week in current clinical care.
Subspecialty depth in Addiction, Child and Adolescent, Forensic, Geriatric, or Consultation-Liaison Psychiatry.
Experience developing OSCEs or serving as an OSCE examiner.
A faculty appointment in a department of psychiatry.
Why This Work Matters
Every major AI system depends on carefully prepared and reviewed human feedback. By applying your psychiatric expertise to model evaluation, you will contribute to the development of AI systems that reason more carefully about complex clinical scenarios and safety-sensitive questions.
OpenTrain makes it possible to build experience in this fast-growing field while taking on flexible, remote project work that fits alongside your professional commitments.
Join OpenTrain to help design taxonomies, rubrics, and benchmarks for detecting self-harm, eating disorders, and suicide risk in AI systems; remote, contractor role (20+ hrs/week) paying USD $50–$90/hr, English required.
Review and rate AI-generated clinical research and biomedical content, create gold‑standard answers, and provide expert written feedback on study design, statistics, and safety; remote contractor role (20+ hrs/week) paying $40–$70 USD/hr and requiring 5+ years clinical research experience.
Join OpenTrain AI as a remote contractor to review and create gold‑standard pharmacy responses that improve AI clinical reasoning; expert pharmacists with PharmD/BPharm and 5+ years’ experience, C1 English, $35–65/hr, 20+ hrs/week. Work includes rating, writing, and RLHF-style evaluation.