Create rigorous scientific programming tasks, verified Python solutions, and tests that evaluate advanced AI models. This freelance project requires strong scientific computing skills and 20+ hours per week.
The Work
You will create and evaluate complex scientific programming tasks used to train and assess advanced AI models. The work combines scientific problem design, Python implementation, testing, model-output review, and quality checks.
- Write scientific problem specifications with one main problem and at least three connected sub-problems.
- Build verified Python solutions with complete unit test coverage.
- Design tests that separate correct from incorrect model outputs and check scientific validity, determinism, and expected results.
- Run structural and quality checks against defined rubrics, then revise tasks using quality feedback.
- Evaluate model-generated outputs for correctness, clarity, and reliability.
- Join review discussions, feedback sessions, and project syncs.
What It Pays and Takes
This is a focused, short-term freelance contractor assignment. The listing does not specify a pay rate.
- Time: 20+ hours per week.
- Work type: Part-time freelance contract.
- Location: Open to contractors in Bangladesh, Brazil, Colombia, Egypt, Ghana, India, Indonesia, Kenya, Nigeria, Pakistan, Turkey, and Vietnam.
- Language: English.
- Education: Master's degree or PhD in a STEM discipline.
- Experience: Intermediate level, with experience in AI data annotation, scientific research, or scientific writing.
- Technical skills: Strong Python programming and scientific computing skills.
- Scientific skills: Ability to write rigorous problems with clear constraints and expected outputs.
- Quality focus: Careful attention to scientific correctness, test discriminativeness, determinism, and task quality.
- Tools and methods: Familiarity with LLM evaluation frameworks or coding benchmarks, plus NumPy, SciPy, SymPy, or comparable scientific tools.
- Background: Published research or academic project experience in a STEM field.
- Helpful experience: Evaluation dataset creation, model-generated code review, scientific programming libraries, or biology and other specialized STEM work.
How It Works
Apply on OpenTrain with your resume, then complete the application on the hiring site.
About AI Training Work
AI training is the human work behind systems that learn from examples, including writing and testing code, rating model responses, and checking whether outputs are correct. OpenTrain helps people build careers in this field, where strong technical and scientific judgment is used to improve AI systems.