I have been doing AI training and data labelling since April 2024
I have been doing AI training and data labelling since April 2024
Hire this AI Trainer
Sign in or create an account to invite AI Trainers to your job.
No software listed
For the past two years, I have specialised in training and evaluating large language models. My work centres on RLHF, Reinforcement Learning from Human Feedback, where I stress test model outputs for factual accuracy, nuance and safety. I don't just rate content I provide the detailed reasoning behind those rating on platforms such as prolific and data annotation. This helps developers identify and fix patterns in model behaviour. I am comfortable navigating complex shifting instructions and pride myself on catching subtle problems and hallucination that other might miss. After thousands of hours of evaluation I have a strong grasp on how human feedback actually shapes and improves AI.
I have been doing AI training and data labelling since April 2024
Degree not specified
2 Years AI training and Data annotation