Use board game strategy, logic, probability, and rule-system expertise to evaluate AI reasoning, create benchmarks, and improve model quality. This remote contractor role requires 20+ hours per week.
Generative AI & RLHF
100% Remote
Worldwide
Eligibility
Entry
Experience
Jul 16, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI recruits contributors for projects that help improve modern artificial intelligence, giving you a place to build a credible profile and grow your experience in this fast-moving field.
Apply in minutes with a free OpenTrain account
Build a portfolio around AI training and evaluation experience
Work remotely on projects connected to cutting-edge AI systems
About AI Training Work
AI training is the human side of building artificial intelligence. Contributors create examples, review model responses, write evaluation criteria, and provide feedback that helps AI systems become more accurate, consistent, and useful.
In this role, your knowledge of board games and strategic systems will support reasoning benchmarks and evaluations. Your judgments can help identify whether an AI understands rules, weighs decisions logically, and explains its conclusions clearly.
Remote work using a computer and internet connection
Flexible contract and part-time structure
Direct contribution to AI reasoning and evaluation quality
The Role
OpenTrain is hiring a Board Game AI Reasoning Evaluator to create and review game-based reasoning tasks that help improve AI systems. You will analyze board game scenarios, strategic decision trees, probability, game mechanics, and complex rule systems.
You will translate tabletop and strategy expertise into prompts, rubrics, benchmarks, and quality standards for AI evaluation. The role is entry level in designation but requires substantial professional or semi-professional experience in board game design, playtesting, tabletop communities, or related strategy-focused environments.
Contractor position
Part-time engagement
Time requirement of 20 or more hours per week
English-language work
Worldwide eligibility
What You'll Do
You will assess both the underlying game logic and the quality of AI-generated responses. Strong work in this role requires careful attention to rules, strategic tradeoffs, probability, and the difference between a plausible answer and a correct one.
Create and review board game reasoning tasks for AI evaluation
Analyze strategic scenarios, decision trees, rule-based systems, and game logic
Evaluate AI-generated answers for correctness, consistency, and reasoning quality
Identify logical errors, rule violations, and flawed reasoning
Develop prompts, rubrics, and evaluation guidelines for strategy-focused tasks
Contribute to benchmark creation and quality assurance for AI evaluation datasets
Collaborate with project teams to improve dataset quality and evaluation methods
Explain strategic decisions clearly in analytical written feedback
Required Qualifications
A bachelor's degree in computer science, mathematics, cognitive science, game design, philosophy, economics, or a related analytical field is required. Candidates should also have at least two years of professional or semi-professional experience in board game design, playtesting, tabletop communities, or related strategy-focused environments.
Strong knowledge of logic, probabilistic reasoning, game mechanics, and complex rule systems
Background in game theory, behavioral economics, decision science, or formal logic
Ability to evaluate AI-generated answers for correctness and rule violations
Strong analytical, problem-solving, and written communication skills
Experience with board game design, playtesting, or tabletop strategy
Ability to create evaluation rubrics, benchmark datasets, or QA frameworks
Helpful Experience
The following experience is helpful but is presented as additional background rather than a substitute for the required analytical and strategy expertise.
Data annotation or AI training
Prompt engineering or quality assurance
Game or puzzle design
Evaluation rubrics or benchmark datasets
Rules-based system analysis
RLHF, model evaluation, or synthetic data generation
LLM benchmarking, Python, SQL, or data analysis tools
Modern board games, trading card games, tabletop role-playing games, strategy games, or competitive game systems
Build Your AI Training Career
AI training and data labeling are among the most accessible ways to participate in the technology behind modern AI. People with specialized knowledge, strong reasoning skills, or careful judgment help shape how models interpret information and respond to complex tasks.
Through OpenTrain, you can turn this project experience into a longer-term AI training portfolio. A stronger profile can help you demonstrate relevant skills, discover matching opportunities, and continue developing in the field.
Apply through OpenTrain with a free account
Showcase strategy, evaluation, and analytical writing experience
Develop experience in prompts, rubrics, benchmarks, and model review
Evaluate AI-generated responses for accuracy, depth, and logical quality while creating expert prompts and reference answers. This worldwide, part-time contractor role offers $245 to $280 per hour for PhD-level expertise.
Use advanced mathematics expertise to evaluate AI-generated solutions, identify subtle errors, and write rigorous exemplars at $80 per hour. This flexible contract role includes paid qualification and project exams.
Use your physics expertise to evaluate AI-generated solutions, identify subtle errors, and write rigorous exemplars for $80 per hour. This worldwide, part-time contractor role requires 17–20 hours weekly and strong scientific English.