Evaluate how well an AI personalization feature uses conversational and account context to produce relevant Japanese responses. Work remotely as an independent contractor for $15 per hour with a 20+ hour weekly commitment.
Generative AI & RLHF
100% Remote Hourly · $15/hr
$15/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Jul 20, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help people discover projects, build a professional AI training profile, and apply in minutes. Creating an OpenTrain account is free.
About AI Training Work
AI training is the human side of building modern artificial intelligence. Contributors review model outputs, write prompts, compare responses, and provide structured feedback that helps AI systems become more useful and reliable.
This work is remote and often flexible, making it possible to build experience in a fast-growing technology field while working from a computer and internet connection.
The Role
OpenTrain is hiring a Japanese AI Personalization Quality Analyst to evaluate a new personalization feature for Gemini. You will assess how effectively the model uses past conversations, Gmail, Search, and YouTube activity to make responses more relevant.
This entry-level contractor position focuses on AI model evaluation and quality analysis. The role is part time, remote, and requires a commitment of 20 or more hours per week. Compensation is $15 per hour.
Employment type: Part-time contractor
Experience level: Entry level
Work arrangement: Remote, worldwide
Language: Japanese
Data type: Text
Labeling activity: Evaluation and rating
What You'll Do
You will use personal experiences and data from your Google account to create realistic evaluation scenarios. You will then review the model's responses and explain which outputs provide the strongest personalization.
Design and execute multi-turn conversational prompts ranging from 1 to 5 turns.
Create prompts using personal context and experiences.
Evaluate responses for grounding, integration, helpfulness, and personalization quality.
Compare and stack-rank two model responses side by side.
Write clear ranking rationales that reference specific conversation turns.
Spot subtle differences in naturalness, including overnarrating.
Maintain strict data hygiene by deleting evaluation conversations.
Requirements
You should be able to read and write Japanese at a high level and communicate your judgments clearly in writing. Strong analytical thinking is important because personalization quality can depend on subtle differences in relevance, naturalness, and context integration.
You will need a desktop or laptop, a reliable internet connection, and a willingness to use a personal Google account for the evaluation work.
High-level Japanese reading and writing proficiency.
Experience designing creative multi-turn prompts using personal context.
Ability to evaluate AI responses for grounding, integration, and helpfulness.
Strong attention to detail when reviewing side-by-side responses.
Excellent written communication for producing clear, structured rationales.
Self-motivation and the ability to work independently in a remote setting.
Desktop or laptop with a good internet connection.
Willingness to use a personal Google account.
Helpful Background
A bachelor's degree or equivalent in a relevant field may be helpful, particularly in linguistics, computer science, or another analytical discipline. Previous experience in data annotation, AI quality evaluation, or content moderation is also useful, but the role is listed at the entry level.
BS or BA degree, or equivalent relevant experience.
Background in linguistics, computer science, or a related analytical field.
Experience with data annotation, AI quality evaluation, or content moderation.
How to Apply
Create a free OpenTrain account, build your profile, and apply in minutes. Highlight your Japanese proficiency, prompt-writing experience, analytical judgment, and ability to provide precise written rationales.
Confirm that you can commit to 20 or more hours per week.
Showcase relevant Japanese-language and AI evaluation experience.
Be prepared to work independently with a personal Google account.
Lead quality review for Japanese AI-generated content and trainer QA work at $55 per hour. Use your Japanese language expertise to improve accuracy, fluency, localization, and rubric-based AI training quality for 20+ hours each week.
Review how effectively AI uses personal context to produce relevant, grounded responses. This remote contractor role offers entry-level applicants flexible work of 20+ hours per week through OpenTrain.
Help improve large language models by analyzing Japanese content, validating claims, and creating challenging training scenarios. This remote contractor role offers flexible 20+ hour scheduling, with 40-hour commitment and US time-zone overlap required.