Skip to content
OpenTrain AIFor AI Companies

Scientific Computing AI Evaluation Expert

Create and evaluate multi-step computational mathematics tasks for AI agents, with expert solutions, tests, grading criteria, and documentation. This remote five-week contractor assignment pays $300 per approved task and requires at least 40 hours per week.

Apply now
OpenTrain AI

Coding & Software

100% Remote Per task · $300/label

$300/label

Compensation

Contract, part-time

Engagement

Remote

Location

Sep 26, 2026

Posted

Open worldwide

The work

You will turn mathematical and research workflows into self-contained, terminal-based tasks that test AI agents. Your work will cover numerical analysis, optimization, statistics, mathematical modeling, probability, dynamical systems, numerical integration, differential equations, matrix computation, stochastic modeling, and algorithm analysis.

  • Build tasks with datasets, equations, model definitions, constraints, and expected outputs.
  • Implement expert solutions in Python, R, Julia, C/C++, Bash, or another relevant language.
  • Create grading criteria covering accuracy, convergence, complexity, feasibility, and mathematical correctness.
  • Define tolerances, stopping criteria, stability requirements, and reproducibility controls.
  • Write automated tests for edge cases and alternative valid implementations.
  • Debug floating-point precision, solver, conditioning, convergence, and performance issues.
  • Document assumptions, mathematical formulations, expected outputs, and known limitations.

What it pays and takes

This is a remote contractor assignment for five weeks. The role requires advanced technical experience in mathematics, statistics, or a closely related field, along with the ability to independently implement, test, and validate computational algorithms.

  • Pay: $300 per approved task.
  • Schedule: At least 6 hours per day and 40 hours per week.
  • Time overlap: At least 4 hours overlapping with Pacific Standard Time.
  • Location: Remote and open worldwide.
  • Language: English.
  • Background: Ph.D., postdoctoral experience, or equivalent advanced technical experience in mathematics, statistics, or a closely related discipline.
  • Programming: Strong ability in Python, R, Julia, C/C++, Bash, or another relevant language.
  • Environment: Experience working in Linux or terminal-based environments.
  • Expertise: Practical experience in numerical methods, optimization, statistics, mathematical modeling, or scientific computation.

How it works

Apply on OpenTrain with your resume and then complete the application on the hiring site.

About AI training work

AI training is the human work behind systems that learn from examples, including writing, testing, and evaluating model outputs. This role uses specialized mathematical and programming knowledge to create reliable evaluations that show whether AI agents can solve computational problems correctly.

Requirements

  • Experience: Entry level
  • Languages: English

How to apply

  1. Apply here on OpenTrain. You create a free account, and we send you to the hiring platform.
  2. Complete your application on the hiring platform.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar open roles

View all AI training jobs

Mobile AI Coding Task Engineer

Create realistic mobile coding tasks, reproducible environments, deterministic checks, and reference solutions for AI models. This remote contractor role takes about 15 hours per week and pays $50 to $100 per hour equivalent, paid per task that meets project specifications.

Coding & Software
Computer Code Programming
Remote · United Arab Emirates, Argentina, Austria +50 more
English
Part-time · Flexible
Entry level
Per task · $50–$100/hr equivalent

Posted Oct 7, 2026

Coding-Agent Benchmark Engineer

Design and evaluate realistic coding-agent tasks using production-like repositories, tests, evaluators, and scoring rubrics. This contract role is open worldwide, requires English fluency and 20+ hours per week, and calls for deep software engineering experience.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level

Posted Oct 7, 2026

Bioinformatics AI Evaluation Task Designer

Design rigorous bioinformatics and computational genomics tasks that test whether AI models can analyze data, write code, and produce verifiable scientific results. This remote five-week contractor assignment pays $150 per approved task.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level
Per task · $150/label

Posted Oct 7, 2026

GenAI Security Evaluation Engineer

Build and test vulnerable GenAI agent and RAG codebases, annotate security risks, and evaluate detection tools at $150 per hour. This contract role requires 20+ hours per week and strong application security experience.

Coding & Software
Computer Code Programming
Remote · Worldwide
English
Part-time · Flexible
Entry level
Hourly · $150/hr

Posted Oct 7, 2026

Backend AI Coding Task Creator

Create realistic backend coding tasks and reliable verifiers for AI systems. This contractor role offers a flexible schedule of about 15 hours per week and pays $30 to $100 per hour equivalent, paid per task that meets project specifications.

Coding & Software
Computer Code Programming
Remote · United Arab Emirates, Argentina, Austria +50 more
English
Part-time · Flexible
Mid-Senior level
Per task · $30–$100/hr equivalent

Posted Oct 6, 2026