Skip to content
OpenTrain AIFor AI Companies

Italian AI Safety Evaluation Expert

Help improve how advanced AI models handle sensitive topics in Italian through prompt writing, structured evaluation, and red-teaming. This worldwide, part-time contract role pays $40 to $44 per hour.

OpenTrain AI

Generative AI & RLHF

100% Remote Hourly · $40–$44/hr

$40–$44/hr

Compensation

Worldwide

Eligibility

Entry

Experience

Sep 4, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover projects, build a professional profile, and apply in minutes.

Creating an OpenTrain account is free, and your work can contribute to a lasting portfolio in a fast-growing technology field.

  • Worldwide opportunity
  • Part-time contractor engagement
  • Less than 20 hours per week
  • Pay of $40 to $44 USD per hour

About AI Safety Evaluation

AI training is the human side of building artificial intelligence. Contributors write examples, evaluate model responses, identify weaknesses, and provide structured feedback that helps AI systems become more useful and safer.

Safety evaluation focuses on how models respond to sensitive or potentially harmful requests. Your Italian language fluency, cultural awareness, and careful reasoning will help assess whether model behavior is appropriate in context.

  • Work remotely with written prompts and conversations
  • Apply structured guidelines to model evaluations
  • Help shape the behavior of advanced AI systems
  • Build experience in AI training and data labeling

The Role

OpenTrain is recruiting an Italian AI Safety Evaluation Expert to assess and strengthen how advanced AI models handle sensitive topics in Italian. The work combines bilingual language fluency, cultural judgment, structured evaluation, and adversarial thinking.

You will evaluate written prompts and conversations, classify content consistently, identify potential safety issues, and document the reasoning behind your decisions. This is an entry-level opportunity, and prior AI or machine-learning experience is not required.

  • Role focus: Italian AI safety evaluation
  • Data type: Text
  • Engagement: Part-time contractor
  • Experience level: Entry level

What You’ll Do

You will use detailed evaluation guidelines to examine prompts, conversations, and model behavior. The role requires sound judgment when handling sensitive and dual-use information, along with the ability to explain classification and safety decisions clearly.

  • Write expert-level prompts in Italian across sensitive subject areas
  • Classify prompts and conversations using structured evaluation guidelines
  • Identify adversarial phrasing and escalation patterns
  • Evaluate model behavior through an Italian linguistic and cultural lens
  • Record clear reasoning for classification and safety judgments
  • Perform red-teaming and evaluation rating tasks on written content

Requirements

This role is suited to a careful Italian-language writer who can apply detailed instructions consistently and reason clearly about sensitive content. A bachelor's degree, completed or in progress, is required.

  • Native or near-native fluency in Italian
  • Business-level written English
  • Ability to write and classify prompts and conversations using structured guidelines
  • Strong written reasoning and careful attention to detail
  • Sound judgment when evaluating sensitive and dual-use information
  • Bachelor's degree completed or in progress

Helpful Background

Previous experience reviewing, grading, or red-teaming written or technical content can be useful. Background in trust and safety, content moderation, policy evaluation, adversarial testing, or Italian linguistic and cultural analysis is also relevant.

AI or machine-learning experience is not required. The most important capabilities are consistent guideline application, thoughtful judgment, and precise written explanations.

  • Content review, grading, or red-teaming experience
  • Trust and safety or content moderation experience
  • Policy evaluation or adversarial testing experience
  • Italian linguistic or cultural analysis experience

Why This Work Matters

Every major AI system depends on people who prepare examples, review outputs, and identify where models need improvement. Safety evaluators play an important role in helping AI systems respond more appropriately across languages and cultural contexts.

Remote AI training work can be a flexible way to build experience in technology. OpenTrain helps you find relevant opportunities, develop your profile, and grow a career in AI training and data labeling.

  • Contribute to the development of safer AI behavior
  • Use Italian language and cultural expertise in a technical setting
  • Work remotely with a flexible part-time schedule
  • Build a portfolio of AI evaluation experience

How to Apply

Create a free OpenTrain account and apply through OpenTrain. Review the role requirements carefully and highlight your Italian fluency, written English, reasoning ability, and any relevant evaluation, moderation, or red-teaming experience.

  • Apply through OpenTrain
  • Confirm your availability for less than 20 hours per week
  • Showcase relevant language, writing, and evaluation skills
  • Begin building your AI training career

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Bilingual Italian Scientific AI Safety Evaluator

Use your scientific expertise and Italian fluency to evaluate AI responses on sensitive chemical, biological, radiological, and nuclear topics. This remote contract role offers 7 hours per week at $50-$54 per hour.

Generative AI & RLHF
Text
Remote · Worldwide
Italian, English
Part-time · Flexible
Entry level
Hourly · $50–$54/hr

Posted Sep 4, 2026

Italian QA Annotator Quality Reviewer

Review and score Italian-language AI outputs and text data for quality, correctness, and guideline alignment. This flexible contract role offers less than 20 hours per week at $17 per hour.

Generative AI & RLHF
Text
Remote · Italy
Italian, English
Part-time · Flexible
Intermediate level
Hourly · $17/hr

Posted Nov 12, 2025

AI Safety LLM Evaluator, French and English

Work remotely as a French and English AI Safety LLM Evaluator, reviewing model responses, red-teaming safety boundaries, and creating evaluation data. Earn $24 to $36 per hour while helping improve safer AI systems.

Generative AI & RLHF
Text
Remote · Worldwide
French
Part-time · Flexible
Intermediate level
Hourly · $24–$36/hr

Posted Apr 3, 2026