Evaluate how conversational AI uses personal context in Portuguese, rank responses, and explain quality issues. This remote contractor role pays $15/hour and requires 20+ hours per week.
Generative AI & RLHF
100% Remote Hourly · $15/hr
$15/hr
Compensation
Worldwide
Eligibility
Intermediate
Experience
Jul 16, 2026
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI is recruiting and contracting for this role, giving contributors a way to build experience in a fast-growing field where people help shape how modern AI systems work.
Creating an OpenTrain account is free. You can build a profile around your AI training experience and apply for opportunities that match your skills.
Remote contractor work available worldwide
Part-time engagement requiring 20+ hours per week
Hourly pay of $15 USD
Contractor employment arrangement
About AI Personalization Evaluation
AI training is the human work behind modern artificial intelligence. Contributors write prompts, review model responses, compare outputs, and provide structured feedback that helps AI systems become more accurate, useful, and natural.
In this role, you will focus on personalization quality: whether a conversational AI system uses relevant personal context appropriately without making unsupported assumptions or forcing connections.
Help evaluate cutting-edge conversational AI
Use human judgment to assess relevance, grounding, and naturalness
Contribute to model improvement through detailed written feedback
Work remotely with a computer and reliable internet connection
The Role
OpenTrain is seeking a Portuguese AI Personalization Quality Analyst to evaluate how conversational AI uses personal context to make responses more relevant and helpful. You will create realistic multi-turn prompts based on your own experiences, review personalized responses, rank alternatives, and explain your judgments in clear written rationales.
This intermediate-level contractor role combines creative prompt design, nuanced AI evaluation, side-by-side response ranking, and structured quality analysis.
Language: Portuguese reading and writing
Subject matter: AI personalization response evaluation
Data type: Text
Workload: 20+ hours per week
Pay: $15 USD per hour
Engagement: Part-time contractor
What You'll Do
You will assess whether conversational AI responses follow the intent of an initial prompt and use personal information appropriately. Your evaluations should be grounded in specific conversation turns and distinguish helpful personalization from inaccurate or unnecessary references to the user.
Design and execute multi-turn conversational prompts, typically spanning one to five turns
Create prompts that require the model to use relevant personal information and experiences
Judge whether responses follow the intent of the starting prompt
Check claims about the user for evidence, accuracy, and flawed inferences
Assess whether personal information is integrated naturally rather than repeated or unnecessarily narrated
Compare model responses side by side
Rank responses for helpfulness, usability, enjoyment, and overall quality
Write concise, defensible rationales tied to specific conversation turns
Explain response strengths, weaknesses, and other quality issues
Required Skills
This role requires strong Portuguese reading and writing ability, excellent analytical judgment, and the ability to evaluate nuanced or ambiguous AI responses. You should be comfortable identifying subtle personalization failures and explaining your decisions clearly.
Reliable computer and internet access, independence, attention to detail, and effective remote communication are also required.
Strong Portuguese reading and writing ability
Analytical judgment when evaluating nuanced personalized AI responses
Ability to identify grounding failures and unsupported inferences
Understanding of incorrect personalization, forced connections, and naturalness
Ability to create creative multi-turn prompts using personal context
Ability to rank responses side by side
Clear, structured written communication
Strong attention to detail and independent work habits
Reliable computer and internet access
Helpful Background
A bachelor's degree or equivalent experience in policy, law, ethics, linguistics, journalism, computer science, or a related analytical field is helpful. Previous experience in data annotation, AI quality evaluation, content moderation, or a related area is preferred.
Bachelor's degree or equivalent experience in a relevant analytical field
Experience with data annotation
Experience with AI quality evaluation
Experience with content moderation or a related area
Why This Work Matters
Every major AI system depends on examples and feedback prepared by people. By designing realistic conversations and evaluating how models handle personal context, you will help improve the quality and usefulness of AI interactions while developing experience in an expanding technology field.
Work on the human side of AI development
Build practical experience in response evaluation and RLHF-style review
Use language skills and analytical judgment in a technology-focused role
Contribute to more accurate, helpful, and natural AI behavior
How to Apply Through OpenTrain
Create a free OpenTrain account, build your profile, and apply for this opportunity through the platform. Highlight your Portuguese proficiency, experience evaluating written content, and ability to provide precise, evidence-based feedback.
Show strong Portuguese reading and writing ability
Describe relevant annotation, AI evaluation, or moderation experience
Emphasize analytical judgment and attention to detail
Demonstrate that you can work independently for 20+ hours per week
Review and improve Portuguese AI responses for accuracy, reasoning, localization, and clarity in a fully remote contractor role based in Brazil. Work 20+ hours weekly at $30 per hour through OpenTrain.
Review and score European Portuguese language data and AI outputs for quality, correctness, and guideline alignment. This flexible contract role offers less than 20 hours per week at $15 per hour.
Use Portuguese fluency and careful judgment to evaluate sensitive AI model behavior, write expert prompts, and identify adversarial language. This remote contract pays $40-$44 per hour for about 7 hours weekly, with training provided.