Evaluate how well AI personalizes Thai-language conversations using creative multi-turn prompts, detailed ratings, and side-by-side response comparisons. This remote three-month contract pays $15 per hour.
Generative AI & RLHF
100% Remote Hourly · $15/hr
$15/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Jul 16, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. Create a free profile, discover projects across the industry, and apply in minutes to work that helps shape the future of artificial intelligence.
About AI Training Work
AI systems improve through human-designed examples and careful evaluations. In this role, your Thai-language prompt design and response analysis will help assess whether an AI system uses personal context accurately, naturally, and helpfully.
Remote work that can fit around your schedule
Hands-on experience evaluating a cutting-edge AI personalization feature
A chance to build experience in prompt design, model evaluation, and data labeling
The Role
OpenTrain is hiring an AI Personalization Quality Analyst for a remote contract focused on Thai-language interactions. You will create multi-turn conversations that require the model to draw on personal information and experiences, then assess the quality of its personalization.
The work combines creative prompt design, nuanced analytical review, and rigorous side-by-side comparison of model responses. The contract lasts three months and pays $15 per hour.
Role level: Entry level
Engagement: Remote contractor role
Schedule: At least 4 hours per day, up to 40 hours per week
Structured time requirement: 20 or more hours per week
Required overlap: 4 hours with Pacific Standard Time
Pay: $15 per hour
What You'll Do
You will evaluate personalization across short, multi-turn conversations and explain your judgments clearly. Your reviews should identify both successful use of context and subtle failures, such as incorrect assumptions or connections that feel forced.
Design and execute multi-turn prompts spanning 1 to 5 turns
Use personal information and experiences from connected personal data sources as conversational context
Evaluate responses for grounding, integration, and helpfulness
Check whether personalization is correct, relevant, and natural
Identify poor inferences, incorrect personalization, and forced connections
Stack-rank two model responses side by side to determine which is more helpful, easy to use, and enjoyable
Write clear, structured ranking rationales that reference specific turn numbers
Delete evaluation conversations to maintain data hygiene and avoid affecting future chat history
Requirements
This role requires high-level Thai reading and writing proficiency, strong analytical judgment, and the ability to communicate detailed evaluations in clear written English or another required work language as applicable. You should be comfortable working independently in a remote environment.
High-level proficiency reading and writing in Thai
Experience designing creative multi-turn AI prompts using personal context
Exceptional analytical thinking when reviewing nuanced or ambiguous AI responses
Strong ability to evaluate grounding, integration, and helpfulness
Ability to identify incorrect personalization, poor inferences, and forced connections
Meticulous attention to detail during side-by-side response review
Excellent written communication for concise, structured rationales
Self-motivation and independence in remote work
Desktop or laptop with a reliable internet connection
Helpful Background
A bachelor's degree or equivalent experience in an analytical field can be helpful. Previous work in AI evaluation, data annotation, or content moderation is also useful, but the role is designated as entry level.
BS or BA, or equivalent experience
Background in policy, law, ethics, linguistics, journalism, computer science, or a related analytical field
Experience with data annotation, AI quality evaluation, or content moderation
How It Works
AI training and data-labeling work is a growing part of the technology industry. Human contributors design prompts, review model outputs, and provide structured feedback that helps AI systems become more useful and reliable.
Create a free OpenTrain account, build your profile around your language and analytical skills, and apply for this project in minutes. If selected, you will complete the assigned evaluation work remotely according to the project schedule.
Apply through OpenTrain
Work remotely with the required Pacific Standard Time overlap
Commit to at least 4 hours per day and up to 40 hours per week
Evaluate how an AI uses personal context to improve responses in Vietnamese. Create multi-turn prompts, rank model outputs, and write detailed rationales in a flexible remote contract.
Review how effectively AI uses personal context to produce relevant, grounded responses. This remote contractor role offers entry-level applicants flexible work of 20+ hours per week through OpenTrain.
Evaluate how well an AI personalization feature uses conversational and account context to produce relevant Japanese responses. Work remotely as an independent contractor for $15 per hour with a 20+ hour weekly commitment.