Review Japanese and English AI-generated content, evaluate safety and reasoning, and provide clear feedback for safer model behavior. This remote contract offers 20+ hours per week at $27-$31 per hour.
Generative AI & RLHF
Remote Hourly · $27–$31/hr
$27–$31/hr
Compensation
1 country
Eligibility
Intermediate
Experience
Apr 3, 2026
Posted
Open to applicants in
Japan
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is the hiring and contracting organization for this role, connecting contributors with meaningful opportunities to help shape how modern AI systems work.
Creating an OpenTrain account is free, and candidates can build a profile and apply in minutes.
Remote contract work for contributors based in Japan
Part-time schedule with a commitment of 20+ hours per week
Hourly pay ranging from $27 to $31 USD
About AI Safety Training
AI training is the human side of building artificial intelligence. Reviewers evaluate model responses, identify risks, and provide feedback that helps AI systems become more accurate, useful, and safe in real-world settings.
In this work, your judgment helps models handle difficult situations responsibly, including harmful requests, adversarial prompts, cultural context, and potentially unsafe content.
Contribute to the development of safer generative AI
Use human expertise to evaluate model behavior and reasoning
Work remotely in a fast-growing field at the cutting edge of technology
The Role
OpenTrain AI is seeking an AI Safety Data Reviewer with near-native or native Japanese proficiency and strong English reading and writing skills. In this remote, hourly-paid contract role, you will review AI-generated content and safety decisions across Japanese and English, assessing reasoning quality, step-by-step problem-solving, accuracy, clarity, and alignment with safety policies.
You may be exposed to explicit, toxic, violent, sexual, or psychologically disturbing material, including content involving sexual or violent topics. This work supports the safe deployment of AI models in real-world environments.
Experience level: Intermediate
Employment type: Contractor and part-time
Location: Japan
Time requirement: 20+ hours per week
Pay: $27-$31 USD per hour
What You'll Do
You will make nuanced, reproducible judgments about AI responses and safety decisions. Your reviews should clearly explain the reasoning behind each rating or comparison and identify how model behavior could be improved.
Review AI-generated content and safety decisions
Evaluate solutions for correctness, clarity, and logical reasoning
Assess step-by-step problem-solving quality
Identify methodological, conceptual, and factual errors
Fact-check responses when needed
Rate or compare multiple responses for safety and policy alignment
Identify edge cases through LLM red-teaming and adversarial testing
Recognize hate and harassment, sexual content, self-harm, violence, bias, illegal goods or services, malicious activities, malicious code, and misinformation
Apply safety standards consistently across Japanese and English content
Account for cultural nuance, slang, coded language, and shifts in context
Provide clear feedback and recommend mitigations for unsafe behavior
Requirements
This role requires advanced bilingual judgment, strong analytical writing, and senior-level experience in safety-related work. You should be comfortable applying detailed standards consistently while preserving meaning, severity, and intent across Japanese and English.
Near-native or native Japanese proficiency in reading and writing
Minimum C1 English proficiency in reading and writing
Bachelor's degree or higher in Communications, Linguistics, Psychology, Law or Policy, Security Studies, or a related field, or equivalent professional experience
Senior-level experience in Trust and Safety, content moderation, policy operations, risk, compliance, investigations, or a related safety function
Proven LLM red-teaming or adversarial testing experience
Strong knowledge of AI safety and content risk domains
Experience applying policy standards across Japanese and English content
Strong analytical writing skills with clear, reproducible rationales
Comfort working with explicit, toxic, violent, sexual, or psychologically disturbing content in a secure remote environment
Localization or translation experience is preferred
Who Should Apply
This opportunity is suited to experienced Trust and Safety professionals, content moderators, policy specialists, risk and compliance practitioners, investigators, and bilingual reviewers who understand how language, culture, and context affect safety decisions.
It may also appeal to professionals with relevant academic or equivalent experience who have demonstrated expertise in adversarial testing, content risk, or responsible AI evaluation.
You can make careful judgments under nuanced or ambiguous conditions
You can explain decisions in concise, evidence-based written rationales
You understand how harmful or misleading content can be disguised through context, slang, or coded language
You are prepared for the emotional demands of reviewing disturbing material
How to Get Started
Create a free OpenTrain account, build your profile, and apply to this opportunity through the platform. If selected, you will complete the contracting and project onboarding steps before beginning remote work.
Apply through OpenTrain
Highlight Japanese and English proficiency
Showcase Trust and Safety, moderation, policy, or red-teaming experience
Explain relevant work with AI safety, adversarial testing, or bilingual content review
Help improve advanced AI models by creating Japanese safety prompts, evaluating sensitive conversations, and identifying adversarial patterns. This worldwide contractor role offers flexible work under 20 hours per week at $48 to $52 per hour.
Use PhD-level chemistry or biology expertise to evaluate Japanese AI responses for scientific accuracy, helpfulness, and safety. This worldwide contract offers flexible part-time work at $68-$72 per hour, with no prior AI experience required.
Lead quality review for Japanese AI-generated content and trainer QA work at $55 per hour. Use your Japanese language expertise to improve accuracy, fluency, localization, and rubric-based AI training quality for 20+ hours each week.