Use deep insurance expertise to evaluate AI-generated underwriting, claims, actuarial, and risk-management work. This remote US contract role pays $60-$80 per hour and requires 8+ years of professional experience.
Generative AI & RLHF
Remote Hourly · $60–$80/hr
$60–$80/hr
Compensation
1 country
Eligibility
Intermediate
Experience
Jul 10, 2026
Posted
Open to applicants in
United States
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. Create a free profile to discover specialized projects, apply in minutes, and build a portfolio that reflects your professional expertise.
Remote AI training and data-labeling opportunities
A profile for showcasing relevant experience and skills
Flexible contract work across a fast-growing industry
About AI Model Evaluation
AI training is the human side of building modern artificial intelligence. Expert contributors review model responses, identify errors, write high-quality examples, and provide structured feedback that helps language models produce more accurate and useful results.
In this role, your insurance judgment will help assess whether AI systems understand practical underwriting, claims, actuarial, and risk-assessment reasoning.
Evaluate AI-generated work against structured standards
Help identify and close gaps in insurance knowledge
Improve the consistency and quality of training data
The Role
OpenTrain is seeking an Insurance AI Model Evaluation Expert to bring practical professional judgment to the development of foundational language models. You will assess insurance-related model outputs, create challenging evaluation tasks, and explain your decisions through precise, defensible written feedback.
This is a remote contractor opportunity for professionals eligible to work in the United States. The role is listed as requiring 20+ hours per week, with an expected schedule of at least 35 hours per week on weekdays.
Contractor and part-time opportunity
Remote work for US-eligible professionals
Compensation of $60 to $80 per hour
Intermediate experience level
What You'll Do
You will work with research and engineering teams and collaborate with other subject-matter experts to maintain reliable, consistent judgments across insurance training data.
Guide teams on underwriting, claims, and risk-assessment reasoning
Design challenging insurance tasks and write accurate solutions
Evaluate model outputs for correctness, judgment, and reasoning quality
Develop and refine insurance-specific evaluation guidance
Create and apply scoring rubrics for AI-generated work
Provide clear feedback grounded in professional insurance practice
Collaborate with experts to maintain consistent evaluations
Requirements
You should have substantial, dedicated experience in insurance and be able to translate nuanced professional judgment into clear explanations and repeatable evaluation standards. Direct experience reviewing AI outputs against rubrics is required.
At least eight years of professional insurance experience
Background in underwriting, claims, actuarial work, or risk management
Hands-on experience evaluating LLM or AI model outputs
Experience using structured rubrics or scoring criteria
Ability to apply practical insurance judgment to risk and claims reasoning
Ability to write accurate solutions and defensible evaluation feedback
Demonstrated progression and increasing responsibility in insurance
Strong written and verbal communication skills
Strong problem-solving and interpersonal skills
Helpful Professional Background
Experience at a recognized insurance organization and direct familiarity with real underwriting or claims decisions are valuable preparation for this work. You should be comfortable turning complex professional reasoning into precise, understandable written guidance.
Real-world underwriting decision experience
Practical claims decision experience
Experience explaining nuanced insurance judgments
Comfort maintaining consistent standards with other experts
How to Apply Through OpenTrain
Create a free OpenTrain account and build a profile highlighting your insurance experience, increasing responsibilities, and AI evaluation background. OpenTrain helps professionals discover and grow careers in AI training and data labeling while connecting their expertise to specialized remote projects.
Apply as a US-eligible professional
Highlight insurance specialization and career progression
Showcase experience with LLM evaluation and structured rubrics
Prepare for a weekday schedule of at least 35 hours per week
Use your professional expertise to evaluate AI-generated responses, apply detailed rubrics, and provide feedback that improves model behavior. This part-time remote contract offers 20+ hours per week and pays $140-$200 per hour.
Use PhD-level physics expertise to adjudicate competing AI model solutions, assess assumptions and approximation limits, and write rigorous evaluations. This worldwide, part-time contract pays $80 to $160 per hour.
Evaluate AI-generated responses for accuracy, depth, and logical quality while creating expert prompts and reference answers. This worldwide, part-time contractor role offers $245 to $280 per hour for PhD-level expertise.