Skip to content
OpenTrain AIFor AI Companies

AI Quality Analyst, Gemini Personalization

Evaluate Gemini personalization through creative multi-turn prompts, detailed response analysis, and side-by-side rankings. This remote contractor role offers flexible hours worldwide for skilled English writers and analytical thinkers.

OpenTrain AI

Generative AI & RLHF

100% Remote

Worldwide

Eligibility

Entry

Experience

Jul 17, 2026

Posted

Open worldwide

Interested in this role?

Create a free OpenTrain account and apply in minutes.

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role and the #1 platform for finding and building careers in AI training and data labeling. OpenTrain helps people start and grow careers teaching AI through flexible projects, professional profiles, and streamlined applications.

  • Apply in minutes through a free OpenTrain account
  • Build experience in a fast-growing AI training industry
  • Work remotely with a computer and reliable internet connection

About AI Training Work

AI systems improve through human feedback, evaluation, and carefully prepared examples. Contributors review model behavior, test difficult scenarios, and explain which responses are more accurate, useful, natural, and aligned with a user’s intent.

  • Help shape how advanced AI systems understand and respond to people
  • Use analytical judgment and written feedback rather than traditional software development
  • Choose flexible work that can fit around studies, another job, or family commitments

The Role

OpenTrain AI is recruiting an AI Quality Analyst to evaluate a new personalization feature for Gemini. In this contractor role, you will assess how effectively the model uses information from past Gemini conversations, Gmail, Google Search, and YouTube activity to make responses more relevant and helpful.

The work combines creative prompt design with rigorous analysis of model outputs. You will test personal-context scenarios, identify incorrect or forced personalization, assess grounding and naturalness, and explain your rankings with clear evidence.

  • Contractor engagement lasting 3 months
  • At least 4 hours per day, up to 40 hours per week
  • At least 4 hours of overlap with PST
  • 20+ hours per week
  • Worldwide opportunity
  • English-language work

What You’ll Do

You will create realistic, multi-turn conversations and evaluate how the model responds to personal context. Reviews should be precise, defensible, and tied to the user’s intent and the evidence available in the conversation or connected activity.

  • Design and execute multi-turn conversational prompts, typically 1 to 5 turns, that require the AI to use personal information and experiences
  • Evaluate whether personalization is appropriately applied to the user’s intent
  • Analyze grounding issues and check whether claims are supported by evidence rather than flawed inferences or hallucinations
  • Assess integration quality, including whether personal data is used naturally without robotic overnarrating
  • Rigorously evaluate and stack-rank two model responses side by side
  • Determine which response is more helpful, easy to use, and enjoyable overall
  • Write clear, structured rationales for rankings and comparisons
  • Reference specific conversation turn numbers in your explanations
  • Identify incorrect personalization, poor inferences, forced connections, and subtle differences in naturalness

Required Qualifications

This is an entry-level opportunity for candidates who can combine strong English comprehension with careful, nuanced judgment. You should be comfortable evaluating ambiguous responses and explaining your reasoning in writing.

  • High English proficiency in reading and writing
  • Exceptional analytical thinking when evaluating nuanced and ambiguous AI responses
  • Experience designing creative multi-turn prompts based on personal context
  • Strong understanding of personalization concepts
  • Meticulous attention to detail when reviewing side-by-side model responses
  • Excellent written communication skills
  • Desktop or laptop with a good internet connection

Helpful Background

The following background is helpful but is not listed as required. Experience reviewing AI outputs or other complex content can help you identify subtle quality differences and communicate consistent judgments.

  • BS, BA, or equivalent experience in policy, law, ethics, linguistics, journalism, computer science, or a related analytical field
  • Experience in data annotation, AI quality evaluation, content moderation, or a related role

How to Apply

Create a free OpenTrain account to build your AI training profile and apply for this contractor opportunity. If selected, you will contribute remotely to hands-on evaluation work that helps improve how AI uses context to assist people.

  • Review the role requirements and engagement schedule
  • Highlight your English writing, analytical evaluation, and prompt-design experience
  • Apply through OpenTrain and begin building experience in AI quality work

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Explore related jobs

View all AI training jobs