Skip to content
OpenTrain AIFor AI Companies

Red Team Lead, Offensive Cybersecurity for AI Evaluation

Lead red-team evaluation for AI security: design offensive-evaluation frameworks, safe proxy tasks, and scoring rubrics. Remote contractor role, 20+ hrs/week, $50–$90/hr; requires 5+ years in offensive security, red teaming, or vulnerability research.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $50–$90/hr

$50–$90/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Jul 3, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help people discover projects, build a unified portfolio, and grow durable freelance careers teaching AI. Creating an OpenTrain account is free.

About AI training and red teaming

AI training (data labeling and human evaluation) is the human side of building intelligent systems. Contributors create, evaluate, and refine the examples and tests that shape model behavior — including offensive and defensive security evaluations that help models resist misuse.

This role focuses on offensive cyber work used to evaluate and harden AI systems: designing safe proxy tasks, scoring attacker sophistication, and documenting findings so models and engineering teams can improve security posture.

The role

OpenTrain is recruiting a Red Team Lead to design and validate offensive-security evaluation work that supports AI training and model evaluation. You will develop cyber-capability taxonomies, create safe proxy attack tasks, produce scoring rubrics, review red-team benchmarks, and write clear technical documentation for both technical and non-technical audiences.

This is a remote contractor, part-time role with a minimum commitment of 20+ hours per week. Work is asynchronous and worldwide; English proficiency is required.

  • Employment type: Contractor, part-time
  • Time requirement: 20+ hours/week (asynchronous collaboration)
  • Location: Remote, worldwide (English required)
  • Pay: $50–$90 USD per hour

What you'll do

  • Build taxonomies for cyber-capability tasks and attack stages to structure evaluation work.
  • Design and validate offensive security evaluation frameworks that measure attack coverage and impact.
  • Create safe proxy tasks that simulate advanced attack vectors while enforcing ethical boundaries.
  • Write scoring rubrics for attack sophistication, coverage, and impact to support consistent ratings.
  • Review and strengthen red-team benchmarks to address evolving security risks and exploit chains.
  • Document methodologies and produce technical write-ups for both technical and non-technical stakeholders.
  • Collaborate asynchronously with project stakeholders and incorporate feedback into deliverables.

Requirements

Candidates must preserve safety and ethics while working on adversarial security tasks. OpenTrain expects rigorous handling of sensitive subject matter and adherence to agreed boundaries and controls.

  • 5+ years hands-on experience in offensive cybersecurity, red teaming, exploit development, or vulnerability research.
  • Strong experience with exploit chains, malware analysis, cloud or application security, or social engineering.
  • Proven ability to design evaluation frameworks and scoring rubrics for security work.
  • Clear written and verbal communication for technical and non-technical audiences.
  • Familiarity with ethical boundaries in adversarial security work; certifications or advanced degrees are valued.
  • Experience building cybersecurity benchmarking or evaluation frameworks is a plus.

Technical details of the engagement

This project involves text-based evaluation tasks used for RLHF and structured evaluation ratings. You will produce frameworks, rubrics, and documentation rather than executing real-world attacks. All work must follow ethical safeguards and project-specific constraints.

  • Data type: Text
  • Labeling work: RLHF, evaluation rating
  • Payment type: Hourly (USD $50–$90/hr)

Who should apply and how it works

Apply if you are an experienced offensive security professional who wants to shape how AI systems are tested and hardened. This is ideal for practitioners who enjoy writing clear methods, building benchmarks, and collaborating remotely on safety-focused work.

To apply, create an OpenTrain account, complete your profile, and submit your application for the Red Team Lead role. OpenTrain will manage contracting and project details; selected contributors will work asynchronously and deliver frameworks, tests, and documentation according to project milestones.

  • Great for senior red teamers, exploit developers, and vulnerability researchers who want flexible, remote work.
  • OpenTrain is the hiring organization and manages payments and contracts for this role.

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar Jobs

View all jobs

AI Red Team Engineer — LLM Security

Join OpenTrain AI to red-team large language models and agents, building reproducible attack tests, automation, and security tooling. Part-time remote work for certified cybersecurity professionals with strong scripting and pentesting experience, flexible hours, pay up to $55/hr.

Generative AI & RLHF
Computer Code Programming
Remote · Worldwide
Part-time · Flexible
Intermediate level
Hourly · $40/hr

Posted Nov 17, 2025

AI Red Team Engineer — LLM Security & Pentesting

Part-time, remote contractor role ($40/hr) designing and running adversarial LLM evaluations and RAG pentests. Requires pentesting experience, Python/Bash/PowerShell skills, container/CI-CD security knowledge, C1 English, and immediate HackerRank + platform assessment.

Generative AI & RLHF
Text
Remote · Worldwide
Part-time · Flexible
Intermediate level
Hourly · $40/hr

Posted Oct 6, 2025

AI Safety Red Team Specialist (English & Bengali)

Join OpenTrain to red team conversational AI: probe models with adversarial prompts, document reproducible attacks, and help improve safety. Part-time contractor role (20+ hrs/week), $20–22/hr, requires expert red teaming experience and fluency in English and Bengali.

Generative AI & RLHF
Text
Remote · Worldwide
English, Bangla
Part-time · Flexible
Expert level
Hourly · $20–$22/hr

Posted Jul 13, 2026