Work remotely as a Hebrew–English bilingual evaluator, reviewing and improving AI-generated text for accuracy, clarity, and reasoning. Part-time contractor role at $32/hr (under 20 hrs/week) for experienced translators, editors, or linguistic QA specialists.
Generative AI & RLHF
100% Remote Hourly · $32/hr
$32/hr
Compensation
Worldwide
Eligibility
Entry
Experience
Dec 26, 2025
Posted
Open worldwide
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain is the #1 platform for people building careers in AI training and data labeling. We hire contributors directly and provide projects that let you gain experience shaping how modern AI systems behave.
This role is staffed by OpenTrain; you'll contract directly with us and work on bilingual tasks that help train and evaluate generative AI systems.
About AI training work
AI training (also called data labeling or human feedback) is the human side of building intelligent systems. Contributors annotate, evaluate, and improve model outputs so AI systems learn reliably from real examples.
These projects are often fully remote and flexible, letting you work part-time on structured tasks that reward attention to detail and language skill.
100% remote: work from anywhere with a computer and internet.
Flexible, part-time work — fit tasks around your schedule.
Accessible entry points for language and QA professionals; specialized projects pay for domain expertise.
The role
You will evaluate and improve AI-generated Hebrew and English text, focusing on accuracy, clarity, reasoning quality, and adherence to prompts and guidelines. This is a contractor, part-time role for under 20 hours per week at $32/hour.
Work is asynchronous and remote; you will use structured rubrics and an error taxonomy to rate, rank, and revise model outputs, and write clear explanations and model solutions when required.
Employment type: Contractor, part-time.
Time requirement: Less than 20 hours/week.
Pay: $32 USD per hour.
Data type: Text; label types include evaluation/rating, RLHF, and translation/localization.
What you'll do
Review AI-generated Hebrew and English outputs for factual accuracy, linguistic quality, and reasoning.
Identify reasoning gaps, methodological errors, hallucinations, and unclear explanations even when language appears fluent.
Apply rubric-based evaluations and an error taxonomy to rate and rank multiple model responses consistently.
Write high-quality explanations, corrected model solutions, or revised outputs in Hebrew and English.
Fact-check claims when required, using reputable sources to validate information.
Provide clear, actionable feedback that helps developers improve model behavior and dataset quality.
Requirements
Native or near-native Hebrew proficiency with strong writing/editing across formal and informal registers.
English proficiency at C1+ (reading and writing) for bilingual evaluation and guideline adherence.
Bachelor’s degree or higher in Linguistics, Translation, Hebrew Language, Communications, Journalism, or a related field.
3+ years of professional experience in translation/localization, editorial QA, content QA, or linguistic review (or equivalent).
Experience applying linguistic QA rubrics or an error taxonomy with consistent severity/category judgments.
Strong fact-checking habits and comfort validating claims with reputable sources when required.
Ability to identify reasoning gaps, methodological errors, and unclear explanations (beyond surface language issues).
Comfortable working independently in a remote, asynchronous workflow with reliable availability and communication.
Preferences and helpful experience
The following are not required but will help your application stand out.
Familiarity with Israeli cultural context and terminology norms across common domains.
Prior experience with AI data training, annotation, or evaluation workflows.
Experience in translation/localization projects or editorial QA for Hebrew <> English content.
Who should apply
Experienced bilingual language professionals who enjoy close textual analysis, care about reasoning and factual correctness, and can apply structured rubrics consistently should apply.
This role suits translators, localization editors, linguistic QA specialists, and experienced reviewers who want flexible, impactful part-time work shaping how AI handles Hebrew and English.
How it works / next steps
If this sounds like a fit, create a free OpenTrain account and submit your application. We'll review qualifications and may request a short evaluation test to confirm rubric application and bilingual skills.
As a contractor you will receive onboarding with task guidelines, rubrics, and sample annotations. Work is remote and asynchronous; maintain clear communication and reliable availability for the best outcomes.
Join OpenTrain as a remote contractor to evaluate Hebrew and English AI outputs, create bilingual training content, and write model solutions; $50/hr, 20+ hours/week, Israel-based candidates only.
Join OpenTrain AI as a remote, part-time contractor reviewing and red-teaming LLM outputs in Hebrew and English to find safety failures and produce labeled evaluation data. $26–$38/hr, 20+ hours/week; your feedback will directly shape model safety.
Lead Hebrew QA for OpenTrain's AI training projects by reviewing model outputs, providing written feedback, and maintaining style guides. Remote contractor role for candidates based in Israel, 20+ hrs/week at $55/hr.