AI Content Evaluator & Prompt Specialist
Conducted high-quality data annotation and evaluation tasks focused on optimizing large language models (LLMs) through Reinforcement Learning from Human Feedback (RLHF). Evaluated side-by-side AI-generated responses for factual accuracy, logical consistency, technical precision, and strict formatting compliance. Performed text classification, prompt engineering, and response rating across complex scenarios, providing comprehensive, structured justifications for chosen outputs. Adhered to rigorous quality guidelines to eliminate hallucinations and biases, ensuring the generation of safe, reliable, and contextually aware training datasets for machine learning applications.