Evaluate AI-generated health coaching conversations for safety, empathy, accuracy, and evidence-based guidance. This remote, eight-week contractor assignment offers flexible work for health coaching and behavior-change experts.
Medical & Health
100% Remote
Worldwide
Eligibility
Entry
Experience
Aug 9, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help contributors discover specialized projects, build a professional profile, and grow experience in a rapidly expanding field.
About AI Training Work
AI training is the human side of building artificial intelligence. Reviewers evaluate model responses and provide structured feedback so conversational AI can become more accurate, helpful, safe, and aligned with real-world needs.
The Role
OpenTrain is seeking a Health Coaching AI Output Evaluator to assess large language model outputs involving health coaching, behavior change, and wellness. You will review coaching conversations, recommendations, and plans to determine whether they are accurate, safe, empathetic, actionable, and aligned with evidence-based coaching practices.
Your evaluations will support the improvement of conversational AI used for health-related guidance. This is an individual remote contractor assignment lasting eight weeks, with a commitment of 20 or more hours per week.
Engagement type: Freelance contractor and part-time
Duration: Eight weeks
Time commitment: 20+ hours per week
Work arrangement: Remote and worldwide
Working language: English
What You'll Do
You will apply structured judgment to health coaching content and explain your ratings clearly. The work includes identifying both strong coaching practices and risks that could affect a person's health or well-being.
Review AI-generated or human coaching dialogue for accuracy, tone, empathy, and actionability.
Evaluate goal-setting effectiveness, motivational techniques, personalization, and other coaching quality dimensions.
Identify unsafe, misleading, or non-evidence-based health advice and flag it for correction.
Compare responses with motivational interviewing and behavior-change frameworks.
Provide structured feedback and ratings that support model improvement.
Help refine annotation rubrics when recurring quality issues emerge.
Requirements
A background in health coaching, behavior change, public health, nursing, or a related field is required. You should understand coaching methodologies and behavior-change principles, communicate clearly in writing, and exercise sound judgment when reviewing health-related content.
Health coaching or behavior-change expertise
Understanding of coaching methodologies and motivational interviewing
Sound judgment when identifying unsafe or non-evidence-based health guidance
Ability to apply structured evaluation rubrics consistently
Clear written communication
Fluency in English
Preferred Background
A Certified Health Coach credential, such as NBHWC or a similar certification, is preferred. Experience evaluating conversational AI or chatbot outputs and familiarity with motivational interviewing or related counseling techniques are also valuable.
You should be comfortable distinguishing evidence-based recommendations from unsafe or misleading advice and explaining the reasoning behind your ratings.
Certified Health Coach credential, such as NBHWC or similar
Experience evaluating conversational AI or chatbot outputs
Familiarity with motivational interviewing or related counseling techniques
Why This Work Matters
Health-related AI must be evaluated carefully by people who understand coaching quality, behavior change, and safety. By reviewing model outputs, you can help shape how emerging conversational systems communicate guidance and support healthier user experiences.
Work remotely from anywhere in the world
Contribute to cutting-edge AI training
Use your health coaching or behavior-change expertise in a flexible freelance setting
Lead evaluation of AI-generated medical claims and billing outputs to improve claims processing and payer compliance. Remote US-only contractor role at $80/hr, 20+ hours/week for experienced medical billing managers.
Join OpenTrain AI as a US-based Medical Coding AI Evaluation Lead — remote contractor, 20+ hrs/week at $80/hr — evaluate AI-generated ICD-10, CPT/HCPCS, and DRG coding, lead audits, and improve coding accuracy and revenue integrity.
Apply clinical-trial biostatistics expertise to evaluate AI-generated analyses, verify statistical outputs against analysis plans, and create rigorous evaluation examples. Remote contract work pays $60–$65 per hour for 20+ hours weekly.