AI Response Evaluation & RLHF Specialist
Worked on AI training and evaluation projects involving reinforcement learning from human feedback (RLHF), response ranking, prompt analysis, and quality assessment of AI-generated content. Evaluated model outputs for factual accuracy, reasoning quality, instruction adherence, safety, and overall effectiveness according to detailed annotation guidelines. Performed comparative assessments between multiple AI responses, identified hallucinations and logical inconsistencies, and provided structured feedback to improve model performance. Worked with multilingual content (English and Spanish) across technical and non-technical domains while maintaining high quality standards, consistency, and attention to detail.