Help advance large language models by designing challenging physics problems, writing rigorous solutions, and shaping evaluation benchmarks. This expert-level remote contract offers 20+ hours per week for graduate-level STEM specialists.
Generative AI & RLHF
100% Remote
Worldwide
Eligibility
Expert
Experience
Jul 17, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping contributors discover projects, build a professional profile, and apply in minutes.
Creating an OpenTrain account is free, and this opportunity is available as a remote contractor role worldwide.
About AI Training and LLM Evaluation
AI training is the human side of building artificial intelligence. People create examples, review model outputs, write feedback, and design evaluations that help modern AI systems become more accurate, useful, and reliable.
In this role, your physics expertise will help test how well large language models handle abstraction, symbolic manipulation, multi-step reasoning, and advanced scientific concepts.
Remote work that can fit around other commitments
Directly contribute to the development and evaluation of cutting-edge AI systems
Use advanced subject knowledge in a specialized AI training project
The Physics LLM Evaluation Expert Role
OpenTrain is recruiting an expert in physics and related STEM fields to help fine-tune and evaluate large language models. You will create difficult problems that test model limits, produce high-quality step-by-step solutions, and collaborate with LLM researchers on evaluation goals.
The work spans physics curricula from early undergraduate through PhD-level topics. It emphasizes rigorous reasoning, clear communication, abstraction, symbolic manipulation, and the design of new evaluation benchmarks.
Role type: Part-time contractor
Time requirement: 20+ hours per week
Work arrangement: Remote and worldwide
Language: English
Experience level: Expert
Data focus: Text generation and evaluation rating
What You’ll Do
You will turn advanced physics knowledge into precise, challenging evaluation material. Your work will help researchers understand where language models succeed, where they fail, and how their reasoning can be assessed more effectively.
Design and solve challenging STEM problems that probe model limitations.
Write clear, detailed, step-by-step solutions with well-articulated reasoning.
Collaborate with LLM researchers to align problems with evaluation goals.
Help define new evaluation benchmarks based on physics curricula.
Break down complex STEM concepts into simple, precise explanations.
Provide detailed annotations and constructive feedback.
Required Qualifications
This position is intended for an experienced STEM specialist who can reason carefully through advanced physics problems and communicate solutions in structured English. Your background should be consistent with graduate, PhD, or postdoctoral STEM study or work.
Strong analytical and research ability
Solid STEM background, especially in physics, mathematics, and related sciences
Graduate, PhD, or postdoctoral STEM background
Excellent English comprehension and structured communication
Ability to write clear step-by-step solutions
Experience with abstraction, multi-step reasoning, or symbolic manipulation
Ability to work independently in a remote setting
Helpful Background
Applied physics or a related STEM specialization is helpful. You should also be comfortable explaining difficult physics concepts in simple language and using clear visuals and reasoning when appropriate.
Comfort with abstract reasoning and symbolic manipulation
Experience developing or assessing challenging problem-solving tasks
How to Apply Through OpenTrain
Create a free OpenTrain account, build your profile around your physics and STEM expertise, and apply for this contract opportunity in minutes. If selected, you will contribute to text-based LLM evaluation work as part of a flexible, remote AI training career.
Create or update your free OpenTrain profile
Highlight graduate, PhD, or postdoctoral STEM experience
Showcase physics reasoning, communication, and evaluation strengths
Apply for the role and follow the project instructions
Use advanced physics knowledge to design, solve, and evaluate challenging problems for large language models. This remote, part-time contract role offers 20+ hours per week and welcomes candidates worldwide.
Use PhD-level physics expertise to adjudicate competing AI model solutions, assess assumptions and approximation limits, and write rigorous evaluations. This worldwide, part-time contract pays $80 to $160 per hour.
Create challenging mathematics problems and rigorous solutions that reveal how large language models handle abstraction, symbolic manipulation, and multi-step reasoning. This remote US freelance role offers 30- or 40-hour weekly commitments.