Evaluate how AI uses conversation history and personal data to create helpful, natural responses in Indonesian. This remote contract pays $15 per hour and requires 20+ hours weekly.
Generative AI & RLHF
100% Remote Hourly · $15/hr
$15/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Jul 16, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role and the #1 platform for finding and building careers in AI training and data labeling. Create a free profile to showcase your experience, discover relevant projects, and apply in minutes.
In this role, you will contribute to the quality of AI systems by reviewing how effectively they use personal context while avoiding unsupported or unnatural conclusions.
About AI Training Work
AI training is the human side of building artificial intelligence. Contributors prepare examples, evaluate model outputs, and provide structured feedback that helps AI systems become more accurate, useful, and natural.
This is remote, flexible work in a fast-growing technology field. Your evaluations will help shape how personalized AI features understand context and respond to people.
The Role
OpenTrain is seeking an Indonesian Personalized AI Response Evaluator to assess how well an AI personalization feature uses conversation history and connected personal data. You will create realistic prompts based on personal experiences, review model responses, compare outputs, and provide clear feedback for quality improvement.
This is an entry-level, part-time contractor opportunity available worldwide. The expected commitment is 20 or more hours per week, and the rate is $15 per hour.
Role: Indonesian Personalized AI Response Evaluator
Work arrangement: Remote and worldwide
Engagement: Part-time contractor
Experience level: Entry level
Time requirement: 20+ hours per week
Pay: $15 per hour
Primary language: Indonesian
What You'll Do
You will test personalized AI interactions through realistic, multi-turn conversations and assess whether the system uses personal context appropriately. Your written judgments should be specific, concise, and defensible.
Design and execute multi-turn conversational prompts, typically spanning one to five turns.
Create prompts grounded in personal experiences and relevant personal context.
Evaluate whether responses follow the intent of the starting prompt.
Assess whether personalization is relevant, appropriately integrated, and helpful.
Compare two responses for helpfulness, ease of use, naturalness, and overall quality.
Write comparison rationales that reference specific points in the conversation.
Extract and verify debug information showing whether summaries and personal data sources were used correctly.
Delete evaluation conversations to maintain clean personal chat history and data hygiene.
Requirements
You should be comfortable evaluating nuanced and sometimes ambiguous AI responses in Indonesian. The work requires careful judgment, strong written communication, and the ability to explain why one response performs better than another.
High proficiency reading and writing Indonesian.
Strong analytical thinking and attention to detail.
Ability to design creative multi-turn prompts grounded in personal context.
Ability to recognize incorrect personalization, poor inferences, and forced connections.
Skill in comparing AI responses for helpfulness, naturalness, and integration.
Ability to write structured rationales that reference specific conversation turns.
A desktop or laptop with a reliable internet connection.
Ability to work independently.
Helpful Background
A bachelor's degree or equivalent experience in policy, law, ethics, linguistics, journalism, computer science, or a related analytical field is helpful but not required. Experience with data annotation, AI quality evaluation, content moderation, or related work is preferred.
How It Works
Apply through OpenTrain and build a profile that reflects your AI training experience. As you contribute to projects like this one, your profile can help you demonstrate credible skills and discover future opportunities in AI evaluation and data labeling.
Create a free OpenTrain account.
Review the role requirements and submit your application.
Complete assigned evaluation work remotely and independently.
Grow a portfolio of experience helping improve modern AI systems.
Evaluate Indonesian AI model prompts and conversations for safety, adversarial phrasing, and escalation patterns. This part-time remote contractor role pays $18-$22 per hour and welcomes applicants worldwide.
Help improve personalized conversational AI by creating Dutch prompts, comparing responses, detecting unsupported personalization, and writing evidence-based evaluations. This remote contractor project pays $20 per hour for at least 20 hours weekly.
Review AI-generated responses using email and business application context, assess personalization and relevance, and provide structured feedback. This US-based contract role offers 20+ hours per week for careful analytical evaluators.