Use doctoral-level philosophy expertise to create and evaluate rigorous AI benchmark questions in a worldwide, remote contract role paying $50 to $63 per hour. Work 20+ hours weekly through OpenTrain.
Generative AI & RLHF
100% Remote Hourly · $50–$63/hr
$50–$63/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Aug 28, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is recruiting for this contract opportunity and helps specialists showcase their expertise, discover relevant projects, and build a lasting AI training portfolio.
Creating an OpenTrain account is free, giving you a central place to apply for opportunities and grow your experience in this rapidly developing field.
Worldwide opportunity
Remote contract work
Part-time schedule of 20+ hours per week
Pay of $50 to $63 per hour
About AI Training and Benchmark Assessment
AI training is the human work behind modern artificial intelligence. Experts write examples, evaluate model outputs, and create carefully designed assessments that help AI systems improve their reasoning, accuracy, and reliability.
In this role, your philosophy expertise will support benchmark development by testing deep conceptual understanding across areas including formal ontology and knowledge representation, AI ethics, applied epistemology, philosophy of technology and robotics, and philosophy of science.
Contribute to cutting-edge AI research and evaluation
Apply subject-matter expertise to model assessment
Work remotely with flexible, part-time contract hours
The Role
OpenTrain is seeking a Philosophy AI Assessment Specialist to create and review academic assessment content for an AI research initiative. You will combine advanced subject knowledge with careful judgment about accuracy, clarity, rigor, difficulty, and solution quality.
The work involves producing challenging multiple-choice questions and evaluating pre-written questions to ensure they are precise, self-contained, unambiguous, and solvable.
Role: Philosophy AI Assessment Specialist
Data type: Text
Work areas: Text generation, question answering, and evaluation or rating
Experience level listed for the project: Entry level
What You'll Do
You will author original questions grounded in your philosophy specialization and review existing questions for academic and assessment quality. Each item should test meaningful conceptual understanding while giving a careful reader enough information to determine the answer.
You will also document necessary edits, calibrate difficulty, explain solutions, and support each question with reputable academic sources.
Review questions for accuracy, clarity, completeness, precision, and solvability
Edit and document questions when improvements are needed
Ensure every question is unambiguous, self-contained, and precisely defined
Rate questions as medium, hard, or expert based on conceptual difficulty
Provide one correct answer and nine plausible but subtly incorrect alternatives
Write clear, concise, step-by-step solutions
Include one to five reputable academic references for each question
Requirements
A PhD or doctoral candidacy in Philosophy or a closely related field is required. Candidates should bring strong philosophical argumentation and formal logic skills, knowledge of canonical texts across traditions, and the ability to express complex ideas clearly and concisely in written English.
You must have depth of knowledge in at least one covered area, such as AI ethics, applied epistemology, formal ontology, philosophy of technology, or philosophy of science. Academic citation practice and the ability to assess difficulty, solvability, and solution rigor are also essential.
PhD or doctoral-candidate-level knowledge in philosophy or a closely related field
Strong philosophical argumentation and formal logic skills
Expertise in at least one covered philosophy area
Ability to construct unambiguous, self-contained assessments
Ability to evaluate answer solvability, difficulty, and solution rigor
Excellent written English
Ability to provide reputable academic citations
Helpful Background
Research publications or teaching experience in philosophy are valuable for this work. A master's degree may be considered when paired with exceptional depth in a specific philosophical subdomain.
Philosophy research publications
Philosophy teaching experience
Exceptional master's-level expertise in a specialized subdomain
Work Details and Application
This is a worldwide, remote contractor opportunity with part-time scheduling of 20+ hours per week. The listed hourly pay range is $50 to $63 USD.
AI training and data-labeling work lets specialists contribute directly to how advanced AI systems behave. Create a free OpenTrain account to apply and build a profile that reflects your philosophy and AI assessment experience.
Lead quality assurance for philosophy-focused AI training projects, reviewing arguments, explanations, and trainer work for accuracy and rigor. This remote US contractor role offers up to $65/hour and requires 20+ hours weekly.
Use doctoral-level psychology expertise to create and evaluate rigorous multiple-choice benchmark content for AI research. This fully remote contractor role pays $50-$63 per hour and requires at least 10 hours weekly.
Evaluate AI-generated responses for accuracy, depth, and logical quality while creating expert prompts and reference answers. This worldwide, part-time contractor role offers $245 to $280 per hour for PhD-level expertise.