Skip to content
OpenTrain AIFor AI Companies

Mathematics Quality Assurance Lead

Lead QA for mathematics AI training: evaluate AI-generated math content, coach trainers and QAs, maintain rubrics and onboarding, and help improve QA workflows. Remote US contractor role, 20+ hrs/week, hourly pay up to $75.

OpenTrain AI

Generative AI & RLHF

Remote Hourly · $75/hr

$75/hr

Compensation

1 country

Eligibility

Intermediate

Experience

Jul 9, 2026

Posted

Open to applicants in

United States

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the centralized platform where people build careers in AI training and data labeling. Contributors use OpenTrain to find projects, demonstrate proven work, and grow a durable freelance portfolio in the human side of AI.

We focus on making specialized AI training work easier to track, apply for, and turn into long-term freelance opportunities.

About AI training and this role

AI training (also called data labeling or human evaluation) is the human work that teaches models to reason, calculate, and explain. This role sits at the intersection of mathematics and quality leadership: you will ensure mathematical accuracy, clear reasoning, and consistent reviewer behavior across projects.

This position is a remote contractor role focused on reviewing and improving text-based math outputs and trainer/QAs who evaluate those outputs.

The role

OpenTrain is hiring a Mathematics Quality Assurance Lead to oversee quality and consistency across mathematics AI training projects. You will review AI-generated math content and trainer QA work, provide precise written feedback, and help maintain high standards for mathematical correctness and clarity.

You will also identify recurring quality gaps, support onboarding and calibration, maintain documentation and examples, and help activate contributors who are not working consistently.

  • Role type: Remote contractor, part-time
  • Location: United States only
  • Time requirement: 20+ hours/week
  • Pay: Hourly, up to $75/hour
  • Data type: Text; label type: evaluation/rating

What you'll do

  • Review AI-generated math explanations, proofs, derivations, calculations, word-problem solutions, diagrams, and step-by-step reasoning.
  • Check outputs for correctness, logical reasoning, notation quality, formatting, clarity, and adherence to project rubrics.
  • Provide ongoing written feedback to trainers and QAs and escalate recurring or critical issues.
  • Create and maintain style guides, trackers, FAQs, honeypots, calibration tasks, examples, and onboarding materials.
  • Run onboarding and training calls and help improve scalable QA workflows for mathematics projects.
  • Communicate project updates, quality expectations, and math-specific standards across remote teams.

Requirements

You must have strong mathematics reasoning skills and clear written English suitable for giving precise feedback and creating documentation.

The role expects broad familiarity across a wide range of mathematics topics and experience evaluating mathematical solutions or teaching math.

  • Strong background across algebra, geometry, trigonometry, calculus, linear algebra, discrete math, probability, statistics, number theory, combinatorics, differential equations, and proofs.
  • Ability to evaluate mathematical reasoning, find flawed assumptions, and detect calculation or proof errors.
  • Experience in mathematics, teaching, tutoring, research, quantitative analysis, technical writing, problem creation, assessment design, or math-content review.
  • Experience with AI training, rubric-based review, or LLM evaluation is a strong plus.
  • Comfort collaborating in remote technical environments and producing clear written feedback.

Who should apply

This role is a strong fit if you enjoy reading and critiquing mathematical arguments, coaching reviewers, and building repeatable QA processes for text-based math content.

Apply if you want to shape how AI systems handle mathematical problems and help ensure outputs meet rigorous standards for correctness and clarity.

Schedule, pay, and logistics

This is a U.S.-based remote contractor role and requires at least 20 hours per week. The position is part-time and paid hourly, up to $75/hour.

Work focuses on evaluating text outputs and managing QA activities; familiarity with rubric-based evaluation and remote collaboration tools will help you succeed.

  • Employment types: Contractor, Part-time
  • Languages: English required

How to apply

Create an OpenTrain account (free) and submit an application for this Mathematics Quality Assurance Lead role. Share examples of math review, teaching, or QA work that demonstrate your ability to evaluate proofs, calculations, and mathematical reasoning.

Applications that include brief samples of written feedback or notes on common math errors you spot are especially helpful.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar Jobs

View all jobs

Astronomy Quality Assurance Lead

Lead QA for astronomy and astrophysics AI training: review model outputs and trainer submissions for scientific, mathematical, and editorial accuracy. Remote US-only contractor role, 20+ hours/week, up to $110/hour.

Generative AI & RLHF
Text
Remote · United States
English
Part-time · Flexible
Entry level
Hourly · $110/hr

Posted Jul 8, 2026

Advanced Mathematics LLM Evaluation Expert

Join OpenTrain AI to design and evaluate graduate- and PhD-level mathematics problems that test and improve large language models; work remotely as a contractor for 20+ hours/week building benchmark questions, reviewing model solutions, writing formal Lean proofs, and validating Python computations.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Jul 17, 2026

Mathematics AI Response Evaluation Specialist

Join OpenTrain AI to evaluate and improve AI-generated mathematical answers — $70/hr, 20+ hours/week, remote (selected countries). Ideal for MS/PhD-level mathematicians who can spot errors, write clear model solutions, and rate reasoning quality.

Generative AI & RLHF
Text
Remote · Bangladesh, Bhutan, Brazil +14 more
English
Part-time · Flexible
Entry level
Hourly · $70/hr

Posted Jul 9, 2026