Physics Model Evaluation Expert
Use senior physics judgment to evaluate competing AI model solutions, assumptions, and uncertainty. This remote, part-time expert contract pays $80-$160 per hour and requires a physics PhD.
Posted Aug 3, 2026
Use your professional expertise to evaluate AI-generated responses, apply detailed rubrics, and provide feedback that improves model behavior. This flexible remote contract offers 20+ hours per week and listed rates of $140 to $200 per hour.
Generative AI & RLHF
$140–$200/hr
Compensation
230 countries
Eligibility
Entry
Experience
Aug 4, 2026
Posted
Open to applicants in
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI recruits contributors for specialized projects where human expertise helps shape how advanced AI systems work, and creating an OpenTrain account is free.
This role is part-time remote contractor work. You can build a lasting AI training portfolio while applying your professional knowledge to practical model evaluation tasks.
AI systems learn from carefully reviewed examples, structured judgments, and clear human feedback. In model evaluation work, experts assess whether AI-generated responses are accurate, logical, well-supported, and aligned with professional standards.
Your analysis can help improve model behavior across realistic professional scenarios. Prior AI training experience is helpful, but it is not required when you bring strong real-world domain knowledge and professional writing skills.
OpenTrain AI is seeking an AI Domain Expert for Model Evaluation to support AI training work involving model outputs and professional documents. You will use project rubrics, source materials, and your domain knowledge to make consistent, well-reasoned judgments.
The role is suitable for professionals with experience in software engineering, finance, data science, legal work, or another relevant field. Strong written English and the ability to explain nuanced decisions are central to the work.
You will evaluate model outputs and documentation against project guidelines, recording clear judgments and the evidence behind them. The work may involve reviewing source materials, annotating data, performing quality reviews, and maintaining consistency across repeated tasks.
You will also design and refine prompts based on realistic professional scenarios. Remote collaboration, careful analysis, thoughtful critique, and ethical judgment are important throughout the project.
Candidates must bring relevant professional experience in software engineering, finance, data science, legal work, or another related domain, along with a strong record of producing or reviewing professional documents. The ability to assess AI-generated responses for accuracy, logic, and alignment with domain best practices is essential.
You should be comfortable reviewing or editing complex documents, following detailed project instructions, adapting to evolving requirements, and working independently in a remote environment. Experience with data annotation, prompt engineering, or AI output evaluation is helpful but not essential.
This opportunity is designed for professionals who want to apply specialized knowledge to cutting-edge AI development. It is marked entry level for AI training, so you do not need previous experience in the industry if you can demonstrate strong domain judgment, professional writing, and careful analysis.
The role is available to candidates in the countries listed for this project and requires English fluency. It may be a strong fit for someone seeking flexible, remote work alongside other professional or personal commitments.
Create a free OpenTrain account and apply through OpenTrain AI. Your profile can help you present credible professional experience, discover matching AI training opportunities, and build a portfolio as you contribute to the development of modern AI systems.
Keep exploring
Use senior physics judgment to evaluate competing AI model solutions, assumptions, and uncertainty. This remote, part-time expert contract pays $80-$160 per hour and requires a physics PhD.
Posted Aug 3, 2026
Use your finance expertise to evaluate LLM outputs, identify model weaknesses, and create rubrics and benchmarks for finance-focused AI training. This US contract role offers $100 per hour and requires 20+ hours weekly.
Posted Jul 16, 2026
Evaluate and improve AI models through Python development, response ranking, dataset creation, and RLHF on a fully remote, one-month contractor assignment.
Posted Jul 16, 2026
Browse related job pages
Locations
Languages