Senior AI Trainer & Data Labeling Specialist — Freelance / Remote (Scale AI, Remotasks, Appen)
Served as a Senior AI Trainer and Data Labeling Specialist applying RLHF-style human feedback to train and evaluate large language model responses across GPT and Claude-based systems. Produced quality-controlled annotation outputs and structured feedback for model developers based on bias and consistency checks. Implemented reliable labeling processes through annotation guidelines, SOPs, and reviewer feedback loops. • Trained and evaluated LLM responses using RLHF methodologies • Annotated high-volume NLP items (50,000+ data points across datasets) • Conducted bias audits and quality reviews with structured feedback • Maintained logs and contributed to weekly performance reporting