You will review AI-generated responses to medical and healthcare scenarios. Your judgments will help improve how AI models handle clinical information, reasoning, completeness, and patient safety.
You will write clear, detailed feedback and work with research and project stakeholders to keep evaluation standards aligned.
- Evaluate clinical accuracy and reasoning in AI-generated medical responses.
- Check responses for completeness, unsafe recommendations, gaps, edge cases, and blind spots.
- Apply structured evaluation criteria to patient-safety considerations.
- Write annotations and feedback that help improve AI model behavior.
- Explain complex medical judgments clearly in writing.
- Collaborate with research and project stakeholders on evaluation standards and project goals.
What it pays and takes
This is a contractor engagement for part-time work. The listing is marked entry level, and the role requires active clinical credentials and strong medical judgment.
- Pay: $100 per hour.
- Schedule: Less than 20 hours per week.
- Location: Open worldwide.
- Language: Professional fluency in English.
- Required qualification: Nurse Practitioner or Medical Doctor with active licensure.
- Strong clinical reasoning and knowledge of healthcare workflows.
- Excellent written communication and close attention to detail.
- Helpful background: Medical administration, healthcare operations, or AI annotation.
How it works
Apply on OpenTrain with your resume, then complete the application on the hiring site.
About AI training work
AI training is the human work behind systems that generate and understand information. Healthcare professionals review model responses and explain what is accurate, complete, and safe, helping improve how these systems handle medical questions.