AI Trainer & Data Annotator — Mercor (Handshake.ai) (Remote)
Provided high-quality data annotation, labeling, and preference data for frontier AI model training and fine-tuning. Performed response evaluation, ranking, and RLHF tasks to improve model reasoning, safety, and performance. Applied engineering expertise to support complex technical annotation workflows at scale. • Preference data creation and labeling for fine-tuning. • Response evaluation and ranking for RLHF. • Technical annotation aligned with frontier AI training needs. • Scaled labeling workflows using software engineering practices.