Evaluate personalized AI responses in Indonesian by creating multi-turn prompts, ranking model outputs, and writing clear rationales. This remote contract pays $15 per hour and requires 20+ hours weekly.
Generative AI & RLHF
100% Remote Hourly · $15/hr
$15/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Jul 16, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain helps people find and build careers in AI training and data labeling, with opportunities to contribute to cutting-edge AI systems and grow a professional portfolio.
Creating an OpenTrain account is free, and candidates can build a profile and apply in minutes.
About AI Training and Quality Evaluation
AI training is the human side of building artificial intelligence. People prepare examples, review model behavior, and provide feedback that helps AI systems produce more accurate, useful, and relevant responses.
In this role, your evaluations will focus on how well a personalization feature uses context and responds to Indonesian-language prompts.
Work remotely with a flexible contractor schedule
Help shape the quality of a developing AI personalization experience
Use analytical judgment and written feedback to improve model behavior
The Role
OpenTrain is recruiting an Indonesian AI Quality Analyst to evaluate a new personalization feature for Gemini. You will create short, multi-turn conversations using data from your personal Google account, assess AI responses, compare competing outputs, and explain your decisions in writing.
This is an entry-level, part-time contract role paying $15 per hour. The role description calls for a 30 to 40 hour weekly commitment, while the listing records a minimum availability of 20+ hours per week.
Role: Indonesian AI Quality Analyst, Personalization
Employment type: Contractor and part time
Pay: $15 per hour
Language: Indonesian reading and writing proficiency
Work arrangement: Remote and worldwide
Experience level: Entry level
What You'll Do
You will test the personalization feature through realistic conversations and assess whether the resulting answers use context appropriately. Evaluations require careful side-by-side comparison and concise, well-supported written reasoning.
Design and execute one- to five-turn conversational prompts
Use data from your personal Google account when creating evaluation prompts
Evaluate AI responses for grounding, integration, and helpfulness
Stack-rank two model responses side by side
Write clear rationales explaining your rankings and evaluations
Delete evaluation conversations to maintain strict data hygiene
Requirements
Applicants should be comfortable designing multi-turn conversational prompts based on personal context and judging whether AI responses are grounded, integrated, and helpful. Strong written communication is essential because every comparison must be supported with a clear rationale.
Indonesian proficiency in reading and writing
Creative prompt engineering ability
Strong analytical and evaluative skills for AI responses
Meticulous attention to detail during side-by-side comparisons
Experience designing multi-turn prompts for AI evaluation
Ability to rank AI responses and write clear rationales
BS or BA degree, or equivalent experience, in a relevant analytical field such as Linguistics or Computer Science
Helpful Background
A degree or equivalent background in Policy, Law, Ethics, Linguistics, Journalism, Computer Science, or another relevant field may be helpful. Experience in data annotation, AI quality evaluation, or content moderation is also valuable, but the listing identifies this as an entry-level opportunity.
Apply Through OpenTrain
AI training and data-labeling work offers a way to contribute directly to how modern AI systems behave. Create a free OpenTrain account, build your profile, and apply for this contract opportunity in minutes.
Review the role requirements and weekly availability expectations
Highlight your Indonesian proficiency and evaluation experience
Review how effectively AI uses personal context to produce relevant, grounded responses. This remote contractor role offers entry-level applicants flexible work of 20+ hours per week through OpenTrain.
Evaluate how well AI personalizes Thai-language conversations using creative multi-turn prompts, detailed ratings, and side-by-side response comparisons. This remote three-month contract pays $15 per hour.
Evaluate how well an AI assistant uses personal context in Dutch conversations, compare responses, and explain subtle quality issues. This remote contract role pays $20/hour for 20+ hours per week.