AI Model Training Specialist (Contract) — Mercor AI
Evaluate and rank LLM outputs using structured criteria spanning reasoning quality, factual accuracy, coherence, and domain alignment. Drive RLHF workflows by producing high-quality structured evaluation feedback and surfacing edge-case failure modes. Perform annotation and QA on structured and multimodal datasets (text, audio, video) while following strict guideline compliance. • LLM response evaluation across reasoning, accuracy, coherence, and alignment dimensions • RLHF-oriented feedback production for model training and alignment loops • Hallucination detection and adversarial testing/prompt optimization • Multimodal annotation QA and guideline compliance for text, audio, and video