Skip to content
OpenTrain AIFor AI Companies

Physics Reasoning Evaluator (BS/MS/PhD Required)

Join OpenTrain AI to evaluate and improve physics-focused AI outputs—paid contract work at $80/hr, part-time (minimum ~17–20 hrs/week). Use your physics degree to spot subtle errors, write step-by-step solutions, and rate model responses with detailed rubrics.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $80/hr

$80/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Oct 24, 2025

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We connect subject-matter experts with paid, remote projects that directly shape how modern AI systems learn and behave.

About AI training work

AI training (data labeling/annotation and human feedback) is the human side of building intelligent systems: experts prepare, review, and rate examples that teach models to reason, explain, and solve domain problems. These roles are often remote, flexible, and a direct way to influence state-of-the-art AI.

The role

OpenTrain AI is hiring physics specialists to evaluate AI-generated physics responses. This is a contractor, part-time position focused on advanced physics reasoning and assessment. You will judge correctness, depth of explanation, and clarity while authoring exemplar solutions.

  • Rate type: Evaluation/Rating tasks using detailed rubrics (label type: EVALUATION_RATING).
  • Data may include written derivations and model outputs; project metadata lists video as a supported data type.
  • Employment: Contractor, part-time, fully remote (worldwide).
  • Pay: $80 USD per hour.
  • Time: Minimum ~17–20 hours per week; preferred cadence ~8 hrs/day during active sprints.

What you'll do

Bring deep physics knowledge to evaluate model answers and improve training data quality. Tasks require careful, reproducible judgments and clear, teachable explanations.

  • Review AI-generated physics responses and judge correctness, reasoning depth, and clarity.
  • Identify subtle conceptual, methodological, and computational errors (derivations, assumptions, units, approximations).
  • Fact-check physics claims using reputable public sources and reference precisely when needed.
  • Author exemplar step-by-step solutions and clear explanations using scientific notation and, where appropriate, LaTeX.
  • Rate and compare multiple responses according to detailed evaluation rubrics and provide concise justification for scores.

Requirements

This role is for physics specialists—not generalists. All requirements below are mandatory.

  • BS, MS, or PhD in Physics (or a closely related physics discipline) from a top-100 university (completed or in-progress).
  • Mastery across core areas: classical mechanics, electromagnetism, thermodynamics, quantum mechanics, statistical physics; familiarity with special/general relativity is a plus.
  • Strong quantitative reasoning: dimensional analysis, unit consistency, uncertainty and approximation awareness.
  • Excellent scientific writing: clear, step-wise explanations with correct notation; LaTeX proficiency preferred.
  • Ability to apply and follow detailed evaluation rubrics consistently and with high attention to detail.
  • Availability for a minimum of ~17–20 hours/week; preferred cadence ~8 hours/day during active sprints.

Preferred experience and bonuses

You don't need prior labeling experience to apply if you meet the academic and reasoning requirements, but the items below strengthen your candidacy.

  • Research experience, analytical writing, or formal debate experience.
  • Programming literacy (e.g., Python, MATLAB) to verify computations or reproduce simple results.
  • Prior data labeling, RLHF, or model-evaluation experience is a bonus.

How hiring and onboarding works

OpenTrain AI hires contractors directly. The onboarding process includes paid qualification steps so your time is compensated while we confirm fit and accuracy.

  • Paid qualification exam: 1–2 hours (compensated).
  • Paid project exam: 1–2 hours (compensated).
  • If accepted, you'll receive project briefs, rubrics, and examples; work is delivered remotely on a flexible schedule.
  • Assignments require consistent application of rubrics and clearly documented decisions; feedback and calibration sessions may occur during active sprints.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar Jobs

View all jobs

Senior Physics Research Evaluator (Remote, US, Part-Time)

Join OpenTrain to evaluate frontier physics research and judge AI model reasoning, earning $80–$110/hr. This part-time, remote US role requires a strong publication record and active research experience (PhD/postdoc preferred).

Generative AI & RLHF
Document
Remote · United States
English
Part-time · Flexible
Expert level
Hourly · $80–$110/hr

Posted Jul 10, 2026

Physics LLM Evaluation Expert

Join OpenTrain to design and solve advanced physics problems that probe LLM reasoning and symbolic skills; remote, part-time contractor work (20+ hrs/week) for candidates with graduate-level physics experience and strong written English.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Expert level

Posted Jul 17, 2026

Physics Scientific Reasoning Dataset Engineer

Contractor role designing rigorous scientific-reasoning evaluation datasets for advanced AI models; requires 3+ years of physics experience and 20+ hours/week. Work remotely for OpenTrain AI to author multi-step tasks, reference solutions, and scoring rubrics for model assessment.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 20, 2026