Skip to content
OpenTrain AIFor AI Companies

AI Safety Red Teaming Expert in English and Swedish

Probe conversational AI with jailbreaks, prompt injections, and multi-turn attacks while producing safety data and reproducible vulnerability reports. This remote contractor role supports 20+ hours per week at $48 to $62 per hour.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $48–$62/hr

$48–$62/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Jul 31, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role and the #1 platform for finding and building careers in AI training and data labeling. OpenTrain helps contributors discover meaningful AI work, build a professional profile, and apply in minutes. Creating an OpenTrain account is free.

About AI Safety Training

AI training is the human side of building artificial intelligence. Specialists test and evaluate model behavior, label failures, and provide structured feedback that helps AI systems become more useful and safer. Red teaming is a high-impact form of this work, using adversarial testing to uncover weaknesses before they affect users.

The Role

OpenTrain AI is hiring an AI Safety Red Teaming Expert fluent in English and Swedish. You will probe conversational AI models and agents with adversarial inputs, generate critical safety data, and document vulnerabilities in reproducible reports. This is a remote, text-based contractor role focused on cutting-edge AI safety testing.

  • Contractor position
  • Part-time engagement
  • Worldwide remote opportunity
  • 20+ hours per week
  • Pay: $48 to $62 per hour, with a listed rate of $62 per hour
  • Languages: native English and Swedish fluency
  • Experience level: entry level

What You'll Do

You will conduct structured adversarial testing and turn your findings into high-quality data that can guide improvements to conversational AI systems. Your work will combine creative probing with consistent evaluation, classification, and documentation.

  • Red team conversational AI models and agents with jailbreaks, prompt injections, and multi-turn manipulation.
  • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
  • Follow structured taxonomies, benchmarks, and playbooks to keep testing consistent.
  • Document every test reproducibly.
  • Produce reports, datasets, and attack cases that customers can act on directly.

Required Skills

This role calls for an experienced adversarial thinker who can test systems methodically and communicate findings clearly. You should be comfortable moving between creative attack design, structured frameworks, and practical risk reporting.

  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • A curious and adversarial mindset that instinctively pushes systems to their breaking points.
  • A structured approach using frameworks or benchmarks rather than random testing.
  • Strong communication skills for explaining risks to technical and non-technical stakeholders.
  • Adaptability when moving across projects and customers.
  • Native fluency in English and Swedish.

Helpful Background

Additional experience in any of the following areas can support your work on AI safety testing projects:

  • Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attacks, and model extraction.
  • Cybersecurity, including penetration testing, exploit development, or reverse engineering.
  • Socio-technical risk work, such as harassment or disinformation probing, abuse analysis, or conversational AI testing.
  • Creative probing through psychology, acting, or writing for unconventional adversarial thinking.

Why This Work Matters

Modern AI systems depend on people who can identify unsafe behavior, evaluate model outputs, and prepare reliable training data. By exposing vulnerabilities and describing them precisely, you will contribute directly to the development of safer AI systems while working remotely in a fast-growing technical field.

  • Work at the frontier of AI safety and conversational model evaluation.
  • Use English and Swedish language expertise in adversarial testing.
  • Build experience in AI training, safety evaluation, and structured data production.
  • Work remotely with a flexible part-time schedule of 20+ hours per week.

Apply Through OpenTrain

Create a free OpenTrain account to build your AI training profile and apply for this contractor opportunity. OpenTrain brings together projects across the AI training industry so contributors can find work and grow their careers in one place.

  • Review the role requirements and supported languages.
  • Highlight relevant AI red teaming, cybersecurity, socio-technical, or creative probing experience.
  • Apply through OpenTrain for consideration.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

AI Safety Red Team Expert

Probe conversational AI for jailbreaks, prompt injections, bias exploitation, and manipulation as an expert red team contractor. Work worldwide for $48 to $62 per hour, 20+ hours weekly, using English and Danish.

Generative AI & RLHF
Text
Remote · Worldwide
English, Danish
Part-time · Flexible
Expert level
Hourly · $48–$62/hr

Posted Jul 30, 2026

AI Safety Red Teamer

Challenge frontier AI systems with adversarial prompts, uncover safety weaknesses, and document model behavior across high-risk topics. This expert contractor role pays $70 to $84 per hour for 20+ hours weekly.

Generative AI & RLHF
Text
Remote · United States, Denmark, Estonia +30 more
English
Part-time · Flexible
Expert level
Hourly · $70–$84/hr

Posted Jul 17, 2026

AI Safety Red Team Expert English and Thai

Use expert adversarial testing to expose jailbreaks, prompt injections, bias, misinformation, and harmful behaviors in conversational AI. This remote contractor role offers $24-$35 per hour and about 40 hours per week.

Generative AI & RLHF
Text
Remote · Worldwide
Thai, English
Part-time · Flexible
Expert level
Hourly · $24–$35/hr

Posted Jul 30, 2026