Skip to content
OpenTrain AIFor AI Companies

AI Safety Red Team Evaluator

Design adversarial prompts, rate image-edit outputs against safety rules, and explain model failures in this 1–2 week contractor project. The role requires 20+ hours weekly, fluent English, strong judgment, and reliable internet.

Apply now
OpenTrain AI

Generative AI & RLHF

100% Remote

Contract, part-time

Engagement

Remote

Location

Oct 3, 2026

Posted

Open worldwide

The Work

You will test how large language models handle single-turn image-edit requests. You will classify prompts and outputs, find safety failures, and write clear explanations that can be used to improve model behavior.

  • Classify image-edit prompts and outputs using project guidelines and a defined safety taxonomy.
  • Review ambiguous, borderline, and benign requests consistently.
  • Write precise rationales that support each classification decision.
  • Document bypassed safeguards and subtle policy violations.
  • Create adversarial prompts and compare multiple model outputs.
  • Identify unclear or conflicting guidelines and suggest clarifications.

What It Pays And Takes

This is a project-based independent contractor engagement. The listing does not provide a pay rate. The role is marked entry level, but it requires strong analytical judgment and familiarity with AI safety concepts.

  • Hours: 20 or more hours per week.
  • Term: 1–2 weeks per statement of work.
  • Work arrangement: Part-time, independent contractor, with a self-set schedule.
  • Location: Open worldwide.
  • Language: Fluent English required.
  • Equipment: Your own desktop or laptop and a reliable internet connection.
  • Core skills: Policy-based analysis, careful classification, precise writing, and the ability to assess complex or ambiguous information.
  • Safety experience: Red teaming, prompt engineering, or designing challenge prompts to test AI safety filters.
  • Helpful background: Content moderation, policy analysis, AI safety evaluation, RLHF, or data annotation.
  • Relevant education or experience may include policy, law, ethics, linguistics, journalism, or computer science.

How It Works

Apply on OpenTrain with your resume, then complete the application on the hiring site.

About AI Training Work

AI training work uses human judgments, examples, and written feedback to improve how artificial intelligence systems behave. Safety evaluators are paid for careful policy analysis because their decisions help identify harmful outputs and improve model responses.

Requirements

  • Experience: Entry level
  • Languages: English

How to apply

  1. Apply here on OpenTrain. You create a free account, and we send you to the hiring platform.
  2. Complete your application on the hiring platform.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar open roles

View all AI training jobs

AI Conversation Context Evaluation Expert

Review AI conversations about everyday decisions, identify missing context, and assess whether better questions could improve recommendations. This three-week, 20-plus-hour project is open to qualified candidates in the US, Canada, and Western Europe.

Generative AI & RLHF
Text
Remote · United States, Canada
English
Part-time · Flexible
Entry level

Posted Oct 3, 2026

English-Speaking AI Text Evaluation Writer and Editor

Review and edit AI-generated text for clarity, accuracy, grammar, structure, and tone. This remote contractor role pays $22-$70 per hour and requires at least 20 hours per week.

Generative AI & RLHF
Text
Remote · United Arab Emirates, Argentina, Austria +50 more
English
Part-time · Flexible
Entry level
Hourly · $22–$70/hr

Posted Oct 3, 2026

AI Content Evaluation and Editing Specialist

Review human and AI-generated writing for clarity, accuracy, grammar, tone, and quality. This remote contractor role pays $20 to $70 per hour and requires professional English fluency.

Generative AI & RLHF
Text
Remote · United Arab Emirates, Argentina, Austria +50 more
English
Part-time · Flexible
Entry level
Hourly · $20–$70/hr

Posted Oct 3, 2026

Russian-Speaking Psychology AI Reviewer

Review psychology case studies, assessments, and AI-generated content in Russian and English. This remote contractor role pays $100-$200 per hour and offers flexible, part-time work for candidates with a PhD in psychology.

Generative AI & RLHF
Document
Remote · United Arab Emirates, Argentina, Austria +50 more
Russian, English
Part-time · Flexible
Entry level
Hourly · $100–$200/hr

Posted Oct 2, 2026

Mechanical Design AI Evaluation Engineer

Use advanced mechanical design and CAD expertise to create technical questions, verify answers, and evaluate AI-generated reasoning. This remote contractor role offers $40-$90 per hour for about 15 hours per week.

Generative AI & RLHF
Document
Remote · United Arab Emirates, Argentina, Austria +50 more
English
Part-time · Flexible
Entry level
Hourly · $40–$90/hr

Posted Oct 2, 2026