I have experience evaluating and annotating AI-generated responses for quality, safety, factual accuracy, and adherence
I have experience evaluating and annotating AI-generated responses for quality, safety, factual accuracy, and adherence to guidelines. My work has involved comparing model outputs, identifying strengths and areas for improvement, rating responses across multiple dimensions, and applying RLHF (Reinforcement Learning from Human Feedback) principles to improve model behavior. I am comfortable following detailed annotation guidelines, detecting policy violations, assessing reasoning quality, and providing structured feedback to support the development of reliable and safe AI systems.