Create graduate-level computational challenges that test advanced AI systems using simulations, numerical reasoning, model fitting, and experimental design. This remote contract role pays $70-$90 per hour and requires 20+ hours weekly.
Generative AI & RLHF
100% Remote Hourly · $70–$90/hr
$70–$90/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Aug 13, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this remote AI training role. OpenTrain is the leading platform for finding and building careers in AI training and data labeling, helping contributors discover specialized projects, build a credible profile, and grow their experience in a fast-moving industry.
Creating an OpenTrain account is free, and your profile can help you present relevant expertise as you pursue future AI training opportunities.
Remote contract work with a global, worldwide-eligible opportunity
Part-time structure with a default commitment of 40 hours per week
Compensation of $70-$90 per hour
About AI Training Work
AI training is the human work behind modern artificial intelligence. Specialists create examples, evaluate model behavior, and design rigorous tests that help advanced systems reason more accurately and reliably.
In this role, your statistical and mathematical expertise will shape benchmark tasks that assess whether AI systems can handle authentic scientific and engineering workflows, including difficult edge cases and reproducibility concerns.
Work on cutting-edge AI evaluation and benchmark development
Apply specialized quantitative expertise to challenging model assessments
Contribute remotely with flexible availability around your schedule
The Role
OpenTrain is seeking a Computational Statistics and Applied Mathematics Expert to design and evaluate challenging problems for a large-scale benchmark of advanced AI systems. You will translate authentic scientific and engineering workflows into graduate-level computational challenges involving specialized software, reproducible numerical reasoning, simulation, experimental planning, and interpretation of results.
The strongest problems will test strategic reasoning and awareness of edge cases rather than routine calculations. You will work independently while refining task difficulty, clarity, and reproducibility through testing against advanced AI models.
Remote contractor position
Part-time engagement requiring at least 15-20 hours per week
Default schedule of 40 hours per week
$70-$90 per hour in USD
What You'll Do
You will create original computational problems grounded in statistical, mathematical, or scientific workflows. Each task should provide a meaningful test of advanced reasoning and be sufficiently clear and reproducible for rigorous evaluation.
You will build the supporting Python infrastructure and review model performance, refining tasks when results reveal ambiguity, insufficient difficulty, or important computational edge cases.
Write problem setups, oracle functions, and solution validators in Python
Develop simulations, multi-step numerical computations, model-fitting tasks, and experimental-design problems
Design problems where carefully chosen queries or measurements reveal information hidden in data
Test tasks against advanced AI models
Improve problem difficulty, clarity, and reproducibility
Work across areas such as Bayesian statistics, psychometrics, latent-variable modeling, differential equations, time series, survival analysis, spatial statistics, optimization, numerical linear algebra, and computational geometry
Required Qualifications
You should bring graduate-level expertise in statistics, applied mathematics, a relevant quantitative STEM field, or equivalent research experience. An MS or PhD is required; a PhD is preferred, while an MS should be paired with substantial relevant experience.
You must be able to combine deep quantitative judgment with practical programming and computational skills. Independent work and comfort using remote technical environments are important for completing and validating benchmark problems.
Graduate-level training in statistics, applied mathematics, or a related quantitative STEM field
An MS or PhD, with a PhD preferred or an MS plus substantial relevant experience
Strong Python skills for problem setups, oracle functions, and solution validators
Hands-on proficiency with at least one specialized statistical, mathematical, or scientific software package
Judgment about computational edge cases, reproducibility, and genuine problem difficulty
Comfort working independently in Linux or terminal-based remote compute environments
Helpful Background
Experience in several computational domains or with multiple specialized software packages can be valuable. Deep expertise with one suitable package is more important than familiarity with an entire toolset.
Related experience may come from benchmark or evaluation design, scientific teaching, exam or problem-set development, computational reproducibility, or containerized environments.
Specialized R or Python packages
Matlab or Scilab
statsmodels or PyMC
Benchmark or AI evaluation design
Scientific teaching or assessment development
Computational reproducibility or containerized environments
Who Should Apply
This opportunity is suited to a quantitative researcher, computational statistician, applied mathematician, or related STEM expert who enjoys turning complex real-world reasoning into precise, testable computational problems.
It may be a strong fit if you want to apply advanced technical knowledge to AI development while working remotely and maintaining a flexible contractor schedule.
Quantitative experts who can reason beyond routine calculations
Researchers who care about edge cases and reproducible results
Python-capable specialists comfortable translating theory into executable evaluations
Independent contributors available for at least 20 hours per week
How to Work With OpenTrain
Apply through OpenTrain to be considered for this contract opportunity. If selected, you will contribute to specialized AI training and evaluation work while building experience that can support a longer-term portfolio in the field.
AI training is a rapidly growing area of technology work, and contributors with advanced domain expertise help shape how state-of-the-art systems perform on demanding scientific and technical problems.
Create or update your free OpenTrain profile
Submit your application through OpenTrain
Showcase your quantitative, software, and research experience
Complete approved benchmark development and evaluation work remotely
Use advanced Python and statistical expertise to create and validate challenging research-style problems for AI training. This remote contractor role offers flexible part-time work at $15-$60 per hour.
Use advanced statistical physics expertise to solve, audit, and evaluate complex AI training problems involving disordered systems, quantum codes, and numerical simulations. This remote contractor project pays $80-$160 per hour for approximately 10 hours weekly.
Use advanced statistical expertise to write, review, and evaluate content that helps AI systems improve. This flexible remote contractor role pays $90-$120/hour and is open worldwide.