AI Data Trainer / Prompt Specialist (Freelance)
Refined AI model responses using Reinforcement Learning from Human Feedback (RLHF) workflows to improve output quality. Performed detailed data annotation to support training and iterative performance gains. • Worked on response refinement cycles aligned with human feedback signals. • Annotated or supported annotation processes to improve model behavior. • Focused on improving accuracy and relevance of generated outputs. • Iterated based on evaluation of model response quality.