AI Safety Red Teamer
Test frontier AI systems, uncover jailbreaks and safety failures, and document findings that improve model robustness. This expert contract role offers $70-$84 per hour and requires 20+ hours weekly.
Posted Jul 17, 2026
Test conversational AI models for jailbreaks, prompt injections, misuse, bias, and multi-turn manipulation. This expert contractor role offers $48-$62 per hour and requires native English and Norwegian fluency.
Generative AI & RLHF
$48–$62/hr
Compensation
30 countries
Eligibility
Expert
Experience
Jul 30, 2026
Posted
Open to applicants in
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI hires and contracts contributors for specialized projects where human expertise helps improve the safety, quality, and performance of artificial intelligence.
Create a free OpenTrain account to build your AI training profile and apply for relevant work in minutes. This contractor opportunity is part-time and fully remote within the listed eligible countries.
AI safety red teaming is the process of deliberately probing AI systems to uncover weaknesses before those weaknesses cause harm. Red teamers use adversarial prompts, structured testing, and careful documentation to help teams understand how models respond under pressure.
This work is part of the fast-growing AI training industry, where people evaluate model outputs, create datasets, and provide feedback that shapes how modern AI systems behave. The role may involve sensitive topics including bias, misinformation, and harmful behaviors.
OpenTrain AI is hiring an AI Safety Red Teaming Specialist to probe conversational AI models and agents for jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation. You will surface vulnerabilities, classify failures, flag systemic risks, and produce clear materials that support stronger AI safety coverage.
This is an expert-level role for someone with prior experience in AI red teaming, adversarial machine learning, cybersecurity, or socio-technical probing. The work is text-based and follows clear guidelines, taxonomies, benchmarks, and structured playbooks.
You will test conversational AI systems with realistic and adversarial interactions, then turn your findings into structured evidence. Your reports and datasets should help stakeholders reproduce failures and improve safety evaluations.
This opportunity is intended for an expert contributor who can reason clearly about technical and socio-technical model risks. You should be comfortable working independently within structured testing guidance and communicating findings to both technical and non-technical audiences.
The following experience can strengthen your fit for this project, although it is presented as helpful background rather than a required qualification.
This role is available to contractors located in the eligible countries below. It is not a worldwide opportunity. Candidates apply through OpenTrain, where they can create a free profile and pursue work aligned with their AI training expertise.
Keep exploring
Test frontier AI systems, uncover jailbreaks and safety failures, and document findings that improve model robustness. This expert contract role offers $70-$84 per hour and requires 20+ hours weekly.
Posted Jul 17, 2026
Lead quality assurance for AI red-teaming and safety evaluation projects, reviewing adversarial prompts, risk analyses, and contributor work. This remote U.S. contract role offers up to $100 per hour and requires 20+ hours weekly.
Posted Jul 9, 2026
Use structured red team methods to uncover vulnerabilities in conversational AI systems and agents. This part-time contractor role pays $29-$45 per hour and requires native English and Portuguese fluency.
Posted Jul 30, 2026
Browse related job pages
Expertise
Locations