Help improve how advanced AI models handle sensitive topics in Italian through prompt writing, structured evaluation, and red-teaming. This worldwide, part-time contract role pays $40 to $44 per hour.
Generative AI & RLHF
100% Remote Hourly · $40–$44/hr
$40–$44/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Sep 4, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover projects, build a professional profile, and apply in minutes.
Creating an OpenTrain account is free, and your work can contribute to a lasting portfolio in a fast-growing technology field.
Worldwide opportunity
Part-time contractor engagement
Less than 20 hours per week
Pay of $40 to $44 USD per hour
About AI Safety Evaluation
AI training is the human side of building artificial intelligence. Contributors write examples, evaluate model responses, identify weaknesses, and provide structured feedback that helps AI systems become more useful and safer.
Safety evaluation focuses on how models respond to sensitive or potentially harmful requests. Your Italian language fluency, cultural awareness, and careful reasoning will help assess whether model behavior is appropriate in context.
Work remotely with written prompts and conversations
Apply structured guidelines to model evaluations
Help shape the behavior of advanced AI systems
Build experience in AI training and data labeling
The Role
OpenTrain is recruiting an Italian AI Safety Evaluation Expert to assess and strengthen how advanced AI models handle sensitive topics in Italian. The work combines bilingual language fluency, cultural judgment, structured evaluation, and adversarial thinking.
You will evaluate written prompts and conversations, classify content consistently, identify potential safety issues, and document the reasoning behind your decisions. This is an entry-level opportunity, and prior AI or machine-learning experience is not required.
Role focus: Italian AI safety evaluation
Data type: Text
Engagement: Part-time contractor
Experience level: Entry level
What You’ll Do
You will use detailed evaluation guidelines to examine prompts, conversations, and model behavior. The role requires sound judgment when handling sensitive and dual-use information, along with the ability to explain classification and safety decisions clearly.
Write expert-level prompts in Italian across sensitive subject areas
Classify prompts and conversations using structured evaluation guidelines
Identify adversarial phrasing and escalation patterns
Evaluate model behavior through an Italian linguistic and cultural lens
Record clear reasoning for classification and safety judgments
Perform red-teaming and evaluation rating tasks on written content
Requirements
This role is suited to a careful Italian-language writer who can apply detailed instructions consistently and reason clearly about sensitive content. A bachelor's degree, completed or in progress, is required.
Native or near-native fluency in Italian
Business-level written English
Ability to write and classify prompts and conversations using structured guidelines
Strong written reasoning and careful attention to detail
Sound judgment when evaluating sensitive and dual-use information
Bachelor's degree completed or in progress
Helpful Background
Previous experience reviewing, grading, or red-teaming written or technical content can be useful. Background in trust and safety, content moderation, policy evaluation, adversarial testing, or Italian linguistic and cultural analysis is also relevant.
AI or machine-learning experience is not required. The most important capabilities are consistent guideline application, thoughtful judgment, and precise written explanations.
Content review, grading, or red-teaming experience
Trust and safety or content moderation experience
Policy evaluation or adversarial testing experience
Italian linguistic or cultural analysis experience
Why This Work Matters
Every major AI system depends on people who prepare examples, review outputs, and identify where models need improvement. Safety evaluators play an important role in helping AI systems respond more appropriately across languages and cultural contexts.
Remote AI training work can be a flexible way to build experience in technology. OpenTrain helps you find relevant opportunities, develop your profile, and grow a career in AI training and data labeling.
Contribute to the development of safer AI behavior
Use Italian language and cultural expertise in a technical setting
Work remotely with a flexible part-time schedule
Build a portfolio of AI evaluation experience
How to Apply
Create a free OpenTrain account and apply through OpenTrain. Review the role requirements carefully and highlight your Italian fluency, written English, reasoning ability, and any relevant evaluation, moderation, or red-teaming experience.
Apply through OpenTrain
Confirm your availability for less than 20 hours per week
Showcase relevant language, writing, and evaluation skills
Use your scientific expertise and Italian fluency to evaluate AI responses on sensitive chemical, biological, radiological, and nuclear topics. This remote contract role offers 7 hours per week at $50-$54 per hour.
Review and score Italian-language AI outputs and text data for quality, correctness, and guideline alignment. This flexible contract role offers less than 20 hours per week at $17 per hour.
Work remotely as a French and English AI Safety LLM Evaluator, reviewing model responses, red-teaming safety boundaries, and creating evaluation data. Earn $24 to $36 per hour while helping improve safer AI systems.