CUDA C++ to Python AI Code Evaluation Expert
Evaluate AI-generated code by translating CUDA and C++ implementations into Python with PyTorch and NumPy, reviewing correctness and performance remotely for 20+ hours per week.
Posted Aug 8, 2026
Use strong Python skills to create coding data, evaluate language-model responses, and improve AI training workflows. This fully remote, one-month contractor assignment offers 20, 30, or 40 hours per week.
Coding & Software
Worldwide
Eligibility
Entry
Experience
Jul 16, 2026
Posted
Open worldwide
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is hiring and contracting for this role, giving contributors a way to build experience in a fast-growing field where human expertise helps shape advanced AI systems.
AI training is the human side of building artificial intelligence. Contributors create examples, review model outputs, write evaluations, and provide feedback that helps language models become more accurate, useful, and aligned.
This role focuses on coding and evaluation rather than building or fine-tuning the models themselves. Your Python expertise and technical judgment will help produce reliable training data and assess the quality of model-generated solutions.
OpenTrain is seeking a Python AI Model Evaluation Developer to support large language model improvement through hands-on coding, data generation, and evaluation. You will write Python solutions to code-based questions, create high-quality training examples, compare responses from different models, and provide detailed feedback.
This is a fully remote contractor assignment lasting one month. The role is categorized as entry level, while the stated technical requirements call for at least three years of strong Python programming experience.
You will combine software-development discipline with careful model evaluation. The work includes creating and reviewing technical content, analyzing performance, and delivering feedback that researchers and annotators can use to strengthen AI training processes.
Applicants should bring strong Python programming ability, software-development fundamentals, and the communication skills needed to explain technical judgments clearly. Experience working with model responses and AI training data-generation workflows is also required.
This opportunity is suited to Python developers who enjoy solving technical problems, reviewing code, and making precise quality judgments. It may also appeal to software engineers interested in applying their skills to the rapidly growing field of AI training.
Attention to detail matters because your code, rankings, rationales, and feedback will be used to assess and improve language-model behavior. The assignment requires a consistent weekly commitment and Pacific Time overlap.
Apply through OpenTrain to be considered for this contractor assignment. If selected, you will complete remote AI training and evaluation work within the available 20, 30, or 40 hour weekly commitment.
OpenTrain helps contributors discover and grow careers in AI training and data labeling. Your profile can showcase relevant experience and support a longer-term portfolio as you take on additional opportunities in the field.
Keep exploring
Evaluate AI-generated code by translating CUDA and C++ implementations into Python with PyTorch and NumPy, reviewing correctness and performance remotely for 20+ hours per week.
Posted Aug 8, 2026
Evaluate next-generation AI coding agents by reviewing their workflows, debugging Python outputs, and validating technical results. This remote contractual role requires expert Python engineering experience and practical experience building or using LLM-powered agents.
Posted Aug 15, 2026
Use your Python and AI engineering experience to evaluate coding-agent trajectories, tool calls, code changes, and technical outcomes. Work worldwide on a flexible 20+ hour-per-week contract through OpenTrain.
Posted Aug 30, 2026
Browse related job pages
Expertise
Languages