Evaluate German-language prompts and conversations for AI safety risks, adversarial phrasing, and escalation patterns. This remote contractor role offers flexible work at $48-$52 per hour with training provided.
Generative AI & RLHF
100% Remote Hourly · $48–$52/hr
$48–$52/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Sep 4, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover projects, build a professional profile, and apply in minutes.
Free account creation and profile building
Opportunities to grow experience in a fast-moving AI industry
A place to develop a durable portfolio of AI training work
About AI Training Work
AI training is the human work behind modern artificial intelligence. Contributors write examples, evaluate model responses, classify content, and identify risks so AI systems can become more useful, accurate, and responsible.
This work is remote and often flexible, making it possible to contribute part time from anywhere with an internet connection. Your judgments can directly influence how advanced AI systems respond to sensitive topics.
Remote work from anywhere
Flexible part-time scheduling
Direct contribution to the development of advanced AI systems
The Role
OpenTrain is seeking a German AI Safety Prompt Evaluation Expert to improve how advanced AI models handle sensitive topics in German. You will write and evaluate German-language prompts and conversations, apply structured guidelines, and use cultural judgment to identify potential safety risks.
Prior AI or machine learning experience is not required. Training on the workflow is provided.
Role: German AI Safety Prompt Evaluation Expert
Work type: Remote, part-time contract work
Experience level: Entry level
Data type: Text
What You'll Do
You will review written content carefully and document the reasoning behind your judgments. Success in this role requires consistent application of detailed guidelines, strong written reasoning, and sound judgment when handling sensitive and dual-use information.
Write expert-level prompts in German across sensitive subject areas
Classify prompts and conversations using structured guidelines
Evaluate written content and rate model responses
Identify adversarial phrasing and escalation patterns
Flag potential AI safety risks
Document the reasoning behind evaluation decisions
Maintain consistent quality and careful attention to detail
Requirements
You should be comfortable evaluating sensitive written content in German and explaining your decisions clearly in English. Experience reviewing, grading, or red-teaming written or technical content is helpful but not required.
Native or near-native German fluency
Business-level written English for clear evaluation notes
Bachelor's degree completed or in progress
Strong written reasoning and sound judgment
Ability to assess sensitive and dual-use information responsibly
Ability to follow detailed guidelines consistently
Ability to identify adversarial phrasing and escalation patterns
Careful attention to detail
Schedule, Location, and Pay
This is remote contractor work with a default commitment of about 7 hours per week. Applicants may work from anywhere, with Germany and Western Europe preferred.
Commitment: About 7 hours per week by default
Availability: Less than 20 hours per week
Location: Worldwide, with Germany and Western Europe preferred
Pay: $48-$52 per hour
Employment type: Contractor and part time
Build Your AI Training Career
OpenTrain helps contributors turn individual AI training projects into a stronger professional portfolio. By building your profile and documenting your experience, you can show credible skills, discover relevant opportunities, and grow in the rapidly expanding field of AI training and data labeling.
Apply through OpenTrain
Build a profile around your AI training experience
Grow practical experience in prompt evaluation and AI safety
Evaluate how an AI personalization feature uses information from conversations and digital activity to create helpful German responses. Work remotely as a contractor for $15 per hour, with 20+ hours available weekly.
Use native-level French and careful judgment to evaluate sensitive AI prompts, conversations, and safety behavior. This remote contractor role offers flexible work under 20 hours per week at $48 to $52 per hour.
Work remotely as a French and English AI Safety LLM Evaluator, reviewing model responses, red-teaming safety boundaries, and creating evaluation data. Earn $24 to $36 per hour while helping improve safer AI systems.