Skip to content
OpenTrain AIFor AI Companies

Sports LLM Evaluation Expert

Use deep sports knowledge to write challenging prompts, evaluate large language model responses, and identify factual or reasoning issues. This remote contractor role offers 20+ hours per week for experts with a master's degree and three years of relevant experience.

OpenTrain AI

Generative AI & RLHF

100% Remote

Worldwide

Eligibility

Entry

Experience

Aug 11, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI hires and contracts specialists for projects that help improve the systems behind modern artificial intelligence.

As an OpenTrain contributor, you can build a profile that showcases your experience, discover specialized AI training opportunities, and apply in minutes. Creating an OpenTrain account is free.

About AI Training and LLM Evaluation

Large language models learn and improve through carefully designed examples and human feedback. Experts evaluate model responses, identify errors, and provide clear judgments that help AI systems become more accurate, useful, and reliable.

This work puts subject-matter specialists at the center of cutting-edge AI development. It can be completed remotely and offers flexible project work for people who want to apply their professional knowledge to emerging technology.

The Role

OpenTrain is seeking a Sports LLM Evaluation Expert to assess and improve large language models through rigorous sports-domain review. You will apply knowledge of global sports, leagues, athletes, tournaments, rules, statistics, and analytics to create challenging prompts and evaluate AI-generated responses.

The role involves identifying knowledge gaps, hallucinations, outdated information, logical inconsistencies, and difficult edge cases. You will deliver objective, evidence-based feedback that helps improve model performance and supports high-quality annotation.

  • Role focus: Global Sports LLM Evaluation
  • Work arrangement: Remote and worldwide
  • Engagement: Part-time contractor
  • Time requirement: 20+ hours per week
  • Working language: English
  • Listed experience level: Entry level

What You'll Do

You will create and review sports-focused content used to test language model knowledge, reasoning, accuracy, and response quality. Strong documentation and consistent judgment will be important throughout the work.

  • Create advanced prompts covering global sports, leagues, teams, athletes, tournaments, rules, statistics, and analytics.
  • Evaluate AI-generated responses for factual accuracy, reasoning quality, completeness, nuance, and relevance.
  • Detect hallucinations, logical inconsistencies, outdated information, and challenging edge cases.
  • Develop benchmark datasets and adversarial test cases for sports-related model evaluation.
  • Provide evidence-based feedback supported by reliable references.
  • Collaborate with AI researchers while maintaining annotation quality and clear documentation.

Requirements

This role requires a master's degree or higher and strong knowledge of at least one sport. You should be able to assess sports information accurately and explain your judgments clearly in written English.

At least three years of relevant professional experience is required, preferably in sports journalism, sports media, sports content, research, analysis, reporting, or a related field.

  • Master's degree or higher.
  • Strong knowledge of one or more sports, including relevant leagues, teams, athletes, tournaments, rules, statistics, and key developments.
  • At least three years of relevant professional experience in sports journalism, media, content, research, analysis, reporting, or a related field.
  • Excellent written English and research ability.
  • Strong analytical judgment and attention to detail.
  • Ability to provide accurate, evidence-backed reviews.

Helpful Background

Experience with large language models, generative AI, prompt engineering, or AI evaluation is advantageous. A background in a sports-related academic or professional discipline can also strengthen your application.

  • Sports Management
  • Sports Science
  • Journalism
  • Communications
  • Media Studies
  • Published research, industry recognition, or teaching experience
  • Ability to work independently

Why Work With OpenTrain

OpenTrain helps specialists turn AI training and data-labeling projects into a durable professional portfolio. Your profile can make it easier to show credible experience, find work aligned with your expertise, and grow in a rapidly expanding field.

By evaluating sports knowledge and reasoning, you will contribute directly to how advanced AI systems understand and communicate about the world of sports.

  • Remote work available worldwide.
  • Flexible part-time engagement for 20+ hours per week.
  • Opportunity to apply professional sports expertise to advanced AI systems.
  • Free OpenTrain account and profile-building tools.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Physics LLM Evaluation Expert

Help advance large language models by designing challenging physics problems, writing rigorous solutions, and shaping evaluation benchmarks. This expert-level remote contract offers 20+ hours per week for graduate-level STEM specialists.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Expert level

Posted Jul 17, 2026

Legal LLM Evaluation Analyst

Use your legal reasoning, research, and writing skills to evaluate large language model outputs in a remote, one-month freelance project for contributors in India.

Generative AI & RLHF
Document
Remote · India
English
Part-time · Flexible
Entry level

Posted Aug 7, 2026

Biology LLM Evaluation Expert

Help improve large language models by creating challenging biology problems, writing rigorous solutions, and evaluating model reasoning from undergraduate through PhD level. This remote expert contract requires 20+ hours weekly.

Generative AI & RLHF
Text
Remote · Worldwide
English
Part-time · Flexible
Expert level

Posted Jul 17, 2026