Skip to content
OpenTrain AIFor AI Companies

AI Red Team Expert English And Malay

Probe conversational AI with jailbreaks, prompt injections, and bias tests while documenting vulnerabilities and systemic risks. This remote expert contract requires native English and Malay fluency and offers $17 to $25 per hour.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $17–$25/hr

$17–$25/hr

Compensation

Worldwide

Eligibility

Expert

Experience

Jul 30, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the leading platform for finding and building careers in AI training and data labeling, helping contributors discover projects, build a professional profile, and apply for opportunities in minutes.

Creating an OpenTrain account is free, giving you a practical way to grow experience in a rapidly expanding field where human expertise directly shapes how advanced AI systems behave.

About AI Red Teaming

AI red teaming is part of the human side of building safer artificial intelligence. Experts deliberately test conversational models with challenging inputs, evaluate failures, and produce structured feedback that helps improve model behavior.

This work combines adversarial creativity with careful analysis. Your findings can reveal weaknesses involving bias, misinformation, harmful behavior, prompt manipulation, and other safety risks.

The Role

OpenTrain AI is seeking an AI Red Team Expert fluent in English and Malay to conduct adversarial testing of conversational AI models and agents. You will probe systems, surface vulnerabilities, and generate high-quality red team data that supports AI safety improvements.

The project includes sensitive topics such as bias, misinformation, and harmful behaviors. Participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources.

  • Expert-level contract opportunity
  • Remote work available worldwide
  • Hourly compensation of $17 to $25
  • Default commitment of 40 hours per week, with a listed availability requirement of 20 or more hours weekly
  • Part-time contractor engagement

What You'll Do

You will test conversational AI systems systematically and creatively, then turn observed failures into clear, reproducible evidence. Your work will help technical and non-technical stakeholders understand model weaknesses and prioritize safety improvements.

  • Red team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Annotate failures, classify vulnerabilities, and flag systemic risks
  • Use taxonomies, benchmarks, and playbooks to keep testing consistent
  • Document findings in reports, datasets, and attack cases that stakeholders can act on
  • Probe sensitive areas including bias, misinformation, and harmful behaviors

Required Qualifications

This role requires prior experience with AI adversarial work, cybersecurity, or socio-technical probing, along with the judgment to test systems rigorously and communicate findings clearly.

  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
  • A curious, adversarial mindset and the ability to push systems to breaking points
  • A structured approach using frameworks or benchmarks
  • Strong communication skills for explaining risks to technical and non-technical stakeholders
  • Native fluency in both English and Malay

Helpful Background

Experience in one or more of the following areas can support your work on this project. These backgrounds are helpful rather than additional stated requirements.

  • Adversarial machine learning, including jailbreak datasets, prompt injection, RLHF or DPO attacks, and model extraction
  • Cybersecurity, including penetration testing, exploit development, or reverse engineering
  • Socio-technical risk, including harassment and misinformation probing, abuse analysis, or conversational AI testing
  • Creative probing through psychology, acting, or writing for unconventional adversarial thinking

Remote Schedule And Compensation

This is a remote contractor position open worldwide. The role lists a commitment of 20 or more hours per week, with a default project commitment of 40 hours per week.

  • Work location: Remote, worldwide
  • Engagement: Contractor and part-time
  • Schedule: 20 or more hours per week listed; 40 hours per week default
  • Pay: $17 to $25 per hour

How To Apply Through OpenTrain

Create a free OpenTrain account, build your AI training profile, and apply for this opportunity in minutes. Your experience in adversarial testing, AI safety, cybersecurity, and bilingual communication can help shape the next generation of conversational AI.

  • Highlight English and Malay fluency
  • Describe relevant red teaming, cybersecurity, or socio-technical testing experience
  • Show examples of structured analysis, vulnerability documentation, or model evaluation
  • Review project guidance before participating in sensitive testing work

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

AI Safety Red Team Expert English Indonesian

Probe conversational AI models and agents for vulnerabilities using adversarial testing, jailbreaks, prompt injections, and multi-turn manipulation. This remote, part-time expert contract pays $17 to $25 per hour.

Generative AI & RLHF
Text
Remote · Worldwide
Indonesian, English
Part-time · Flexible
Expert level
Hourly · $17–$25/hr

Posted Jul 30, 2026

AI Safety Red Team Expert English and Thai

Use expert adversarial testing to expose jailbreaks, prompt injections, bias, misinformation, and harmful behaviors in conversational AI. This remote contractor role offers $24-$35 per hour and about 40 hours per week.

Generative AI & RLHF
Text
Remote · Worldwide
Thai, English
Part-time · Flexible
Expert level
Hourly · $24–$35/hr

Posted Jul 30, 2026

AI Safety Red Team Expert

Probe conversational AI for jailbreaks, prompt injections, bias exploitation, and manipulation as an expert red team contractor. Work worldwide for $48 to $62 per hour, 20+ hours weekly, using English and Danish.

Generative AI & RLHF
Text
Remote · Worldwide
English, Danish
Part-time · Flexible
Expert level
Hourly · $48–$62/hr

Posted Jul 30, 2026