Skip to content
OpenTrain AIFor AI Companies

Japanese English AI Safety Data Reviewer

Review Japanese and English AI-generated content, evaluate safety and reasoning, and provide clear feedback for safer model behavior. This remote contract offers 20+ hours per week at $27-$31 per hour.

OpenTrain AI

Generative AI & RLHF

Remote Hourly · $27–$31/hr

$27–$31/hr

Compensation

1 country

Eligibility

Intermediate

Experience

Apr 3, 2026

Posted

Open to applicants in

Japan

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is the hiring and contracting organization for this role, connecting contributors with meaningful opportunities to help shape how modern AI systems work.

Creating an OpenTrain account is free, and candidates can build a profile and apply in minutes.

  • Remote contract work for contributors based in Japan
  • Part-time schedule with a commitment of 20+ hours per week
  • Hourly pay ranging from $27 to $31 USD

About AI Safety Training

AI training is the human side of building artificial intelligence. Reviewers evaluate model responses, identify risks, and provide feedback that helps AI systems become more accurate, useful, and safe in real-world settings.

In this work, your judgment helps models handle difficult situations responsibly, including harmful requests, adversarial prompts, cultural context, and potentially unsafe content.

  • Contribute to the development of safer generative AI
  • Use human expertise to evaluate model behavior and reasoning
  • Work remotely in a fast-growing field at the cutting edge of technology

The Role

OpenTrain AI is seeking an AI Safety Data Reviewer with near-native or native Japanese proficiency and strong English reading and writing skills. In this remote, hourly-paid contract role, you will review AI-generated content and safety decisions across Japanese and English, assessing reasoning quality, step-by-step problem-solving, accuracy, clarity, and alignment with safety policies.

You may be exposed to explicit, toxic, violent, sexual, or psychologically disturbing material, including content involving sexual or violent topics. This work supports the safe deployment of AI models in real-world environments.

  • Experience level: Intermediate
  • Employment type: Contractor and part-time
  • Location: Japan
  • Time requirement: 20+ hours per week
  • Pay: $27-$31 USD per hour

What You'll Do

You will make nuanced, reproducible judgments about AI responses and safety decisions. Your reviews should clearly explain the reasoning behind each rating or comparison and identify how model behavior could be improved.

  • Review AI-generated content and safety decisions
  • Evaluate solutions for correctness, clarity, and logical reasoning
  • Assess step-by-step problem-solving quality
  • Identify methodological, conceptual, and factual errors
  • Fact-check responses when needed
  • Rate or compare multiple responses for safety and policy alignment
  • Identify edge cases through LLM red-teaming and adversarial testing
  • Recognize hate and harassment, sexual content, self-harm, violence, bias, illegal goods or services, malicious activities, malicious code, and misinformation
  • Apply safety standards consistently across Japanese and English content
  • Account for cultural nuance, slang, coded language, and shifts in context
  • Provide clear feedback and recommend mitigations for unsafe behavior

Requirements

This role requires advanced bilingual judgment, strong analytical writing, and senior-level experience in safety-related work. You should be comfortable applying detailed standards consistently while preserving meaning, severity, and intent across Japanese and English.

  • Near-native or native Japanese proficiency in reading and writing
  • Minimum C1 English proficiency in reading and writing
  • Bachelor's degree or higher in Communications, Linguistics, Psychology, Law or Policy, Security Studies, or a related field, or equivalent professional experience
  • Senior-level experience in Trust and Safety, content moderation, policy operations, risk, compliance, investigations, or a related safety function
  • Proven LLM red-teaming or adversarial testing experience
  • Strong knowledge of AI safety and content risk domains
  • Experience applying policy standards across Japanese and English content
  • Strong analytical writing skills with clear, reproducible rationales
  • Comfort working with explicit, toxic, violent, sexual, or psychologically disturbing content in a secure remote environment
  • Localization or translation experience is preferred

Who Should Apply

This opportunity is suited to experienced Trust and Safety professionals, content moderators, policy specialists, risk and compliance practitioners, investigators, and bilingual reviewers who understand how language, culture, and context affect safety decisions.

It may also appeal to professionals with relevant academic or equivalent experience who have demonstrated expertise in adversarial testing, content risk, or responsible AI evaluation.

  • You can make careful judgments under nuanced or ambiguous conditions
  • You can explain decisions in concise, evidence-based written rationales
  • You understand how harmful or misleading content can be disguised through context, slang, or coded language
  • You are prepared for the emotional demands of reviewing disturbing material

How to Get Started

Create a free OpenTrain account, build your profile, and apply to this opportunity through the platform. If selected, you will complete the contracting and project onboarding steps before beginning remote work.

  • Apply through OpenTrain
  • Highlight Japanese and English proficiency
  • Showcase Trust and Safety, moderation, policy, or red-teaming experience
  • Explain relevant work with AI safety, adversarial testing, or bilingual content review

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Japanese AI Safety Evaluation Expert

Help improve advanced AI models by creating Japanese safety prompts, evaluating sensitive conversations, and identifying adversarial patterns. This worldwide contractor role offers flexible work under 20 hours per week at $48 to $52 per hour.

Generative AI & RLHF
Text
Remote · Worldwide
Japanese, English
Part-time · Flexible
Entry level
Hourly · $48–$52/hr

Posted Sep 4, 2026

Japanese Scientific AI Safety Evaluator

Use PhD-level chemistry or biology expertise to evaluate Japanese AI responses for scientific accuracy, helpfulness, and safety. This worldwide contract offers flexible part-time work at $68-$72 per hour, with no prior AI experience required.

Generative AI & RLHF
Text
Remote · Worldwide
Japanese, English
Part-time · Flexible
Entry level
Hourly · $68–$72/hr

Posted Sep 4, 2026

Japanese Content QA Lead

Lead quality review for Japanese AI-generated content and trainer QA work at $55 per hour. Use your Japanese language expertise to improve accuracy, fluency, localization, and rubric-based AI training quality for 20+ hours each week.

Generative AI & RLHF
Text
Remote · Japan
Japanese
Part-time · Flexible
Entry level
Hourly · $55/hr

Posted Jul 9, 2026