AI Quality Analyst, Gemini Personalization
Evaluate Gemini personalization through creative multi-turn prompts, detailed response analysis, and side-by-side rankings. This remote contractor role offers flexible hours worldwide for skilled English writers and analytical thinkers.
Generative AI & RLHF
Worldwide
Eligibility
Entry
Experience
Jul 17, 2026
Posted
Open worldwide
About OpenTrain
OpenTrain AI is the hiring and contracting organization for this role and the #1 platform for finding and building careers in AI training and data labeling. OpenTrain helps people start and grow careers teaching AI through flexible projects, professional profiles, and streamlined applications.
- Apply in minutes through a free OpenTrain account
- Build experience in a fast-growing AI training industry
- Work remotely with a computer and reliable internet connection
About AI Training Work
AI systems improve through human feedback, evaluation, and carefully prepared examples. Contributors review model behavior, test difficult scenarios, and explain which responses are more accurate, useful, natural, and aligned with a user’s intent.
- Help shape how advanced AI systems understand and respond to people
- Use analytical judgment and written feedback rather than traditional software development
- Choose flexible work that can fit around studies, another job, or family commitments
The Role
OpenTrain AI is recruiting an AI Quality Analyst to evaluate a new personalization feature for Gemini. In this contractor role, you will assess how effectively the model uses information from past Gemini conversations, Gmail, Google Search, and YouTube activity to make responses more relevant and helpful.
The work combines creative prompt design with rigorous analysis of model outputs. You will test personal-context scenarios, identify incorrect or forced personalization, assess grounding and naturalness, and explain your rankings with clear evidence.
- Contractor engagement lasting 3 months
- At least 4 hours per day, up to 40 hours per week
- At least 4 hours of overlap with PST
- 20+ hours per week
- Worldwide opportunity
- English-language work
What You’ll Do
You will create realistic, multi-turn conversations and evaluate how the model responds to personal context. Reviews should be precise, defensible, and tied to the user’s intent and the evidence available in the conversation or connected activity.
- Design and execute multi-turn conversational prompts, typically 1 to 5 turns, that require the AI to use personal information and experiences
- Evaluate whether personalization is appropriately applied to the user’s intent
- Analyze grounding issues and check whether claims are supported by evidence rather than flawed inferences or hallucinations
- Assess integration quality, including whether personal data is used naturally without robotic overnarrating
- Rigorously evaluate and stack-rank two model responses side by side
- Determine which response is more helpful, easy to use, and enjoyable overall
- Write clear, structured rationales for rankings and comparisons
- Reference specific conversation turn numbers in your explanations
- Identify incorrect personalization, poor inferences, forced connections, and subtle differences in naturalness
Required Qualifications
This is an entry-level opportunity for candidates who can combine strong English comprehension with careful, nuanced judgment. You should be comfortable evaluating ambiguous responses and explaining your reasoning in writing.
- High English proficiency in reading and writing
- Exceptional analytical thinking when evaluating nuanced and ambiguous AI responses
- Experience designing creative multi-turn prompts based on personal context
- Strong understanding of personalization concepts
- Meticulous attention to detail when reviewing side-by-side model responses
- Excellent written communication skills
- Desktop or laptop with a good internet connection
Helpful Background
The following background is helpful but is not listed as required. Experience reviewing AI outputs or other complex content can help you identify subtle quality differences and communicate consistent judgments.
- BS, BA, or equivalent experience in policy, law, ethics, linguistics, journalism, computer science, or a related analytical field
- Experience in data annotation, AI quality evaluation, content moderation, or a related role
How to Apply
Create a free OpenTrain account to build your AI training profile and apply for this contractor opportunity. If selected, you will contribute remotely to hands-on evaluation work that helps improve how AI uses context to assist people.
- Review the role requirements and engagement schedule
- Highlight your English writing, analytical evaluation, and prompt-design experience
- Apply through OpenTrain and begin building experience in AI quality work
Keep exploring
Explore related jobs
Browse related job pages
Languages