Use political science expertise to create challenging prompts, test large language models, and review responses for accuracy, reasoning, nuance, and current relevance. This eight-week contractor assignment requires 40 hours per week.
Generative AI & RLHF
100% Remote
Worldwide
Eligibility
Entry
Experience
Aug 15, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover projects, build a professional profile, and apply in minutes.
Creating an OpenTrain account is free, and your profile can help you present credible AI training experience as you grow your long-term portfolio.
About AI Training Work
AI training is the human side of building artificial intelligence. People create examples, evaluate model outputs, and provide careful feedback so modern systems can become more accurate, useful, and reliable.
In this role, your political research judgment will directly support the development of large language models. You will work on advanced evaluation material, benchmark datasets, adversarial testing, and evidence-backed reviews.
Contribute to cutting-edge generative AI development
Apply specialized knowledge to practical model evaluation
Work remotely with flexible AI training and data-labeling opportunities across the industry
The Role
OpenTrain is recruiting a Political Science LLM Evaluation Expert to evaluate and improve large language models. You will apply knowledge of political systems, governance, elections, public policy, international relations, and related fields to determine whether model responses are accurate, well-reasoned, complete, nuanced, and current.
The assignment combines political research judgment with hands-on AI training work. You will create demanding evaluation material, identify difficult edge cases, assess annotation quality, and provide objective feedback supported by reliable references.
Category: Generative AI and LLM evaluation
Data type: Text
Languages: English
Work arrangement: Worldwide and remote
Engagement type: Contractor and part time
Experience classification: Entry level in the listing
What You'll Do
You will assess political content across varied topics and help establish clear standards for model performance. Strong documentation and consistent judgment will be important when communicating findings to AI researchers and maintaining annotation quality.
Create advanced prompts covering political science, governance, elections, public policy, and international relations
Evaluate AI-generated responses for factual accuracy, reasoning quality, completeness, nuance, and current relevance
Identify hallucinations, logical inconsistencies, outdated information, and difficult edge cases
Develop benchmark datasets and adversarial test cases
Provide objective feedback supported by reliable references
Collaborate with AI researchers to improve model performance
Maintain high annotation quality and clear documentation
Requirements
A master's degree or higher in any field is required. Degrees in political science, public policy, international relations, government, public administration, law, economics, journalism, and other politics-related disciplines are preferred.
At least three years of relevant experience is expected in political research, public policy, government, political consulting, academia, journalism, international affairs, think tanks, or a related area.
Strong knowledge of political institutions, comparative politics, governance, elections, public policy, and current global political developments
Excellent written English, research, analytical, and documentation skills
Ability to identify factual inconsistencies and evaluate reasoning across varied political topics
Strong attention to detail and ability to produce objective, evidence-backed reviews
Ability to create challenging prompts, benchmark datasets, or adversarial evaluation cases
Helpful Background
Experience with large language models, generative AI, prompt engineering, or AI evaluation is valuable. Published research, teaching experience, or industry recognition is helpful.
This work suits professionals who can operate independently, investigate political claims carefully, and communicate well-supported judgments in written English.
Large language model or generative AI experience
Prompt engineering or AI evaluation experience
Published research or teaching experience
Industry recognition in a politics-related field
Familiarity with reliable research references and evidence-based review
Engagement Details
This is an eight-week contractor assignment scheduled for 40 hours per week, including at least four hours of overlap with Pacific Time. The structured listing also records a time requirement of 20+ hours per week, so confirm the expected schedule during the application process.
Duration: Eight weeks
Scheduled commitment: 40 hours per week
Required overlap: At least four hours with Pacific Time
Contractor assignment
Worldwide eligibility
Build Your AI Training Career With OpenTrain
OpenTrain brings together opportunities for people teaching and evaluating AI, making it easier to build a career in a fast-growing field. Your work reviewing political model outputs can become part of a stronger professional profile and a durable AI training portfolio.
Apply through OpenTrain to take the next step in specialized AI evaluation and help shape how advanced language models understand and respond to political topics.
Help advance large language models by designing challenging physics problems, writing rigorous solutions, and shaping evaluation benchmarks. This expert-level remote contract offers 20+ hours per week for graduate-level STEM specialists.
Use advanced physics knowledge to design challenging problems, solve them step by step, and help evaluate how large language models reason. This flexible remote contractor role is open worldwide.
Assess AI-generated responses for accuracy, logic, relevance, and completeness while creating detailed feedback and training examples. This remote freelance assignment offers flexible work of 20+ hours per week.