RLHF and AI Model Evaluation Contributor
Contributed to AI performance improvement projects using reinforcement learning from human feedback (RLHF). Designed and evaluated prompts to test AI creativity, long-context memory, and decision-making. Performed model evaluations to assess accuracy and compliance with task instructions. • Developed test scenarios requiring logical reasoning skills • Created and rated narrative prompts for AI systems • Evaluated AI-generated responses for quality control • Provided structured feedback to enhance AI model behavior