AI Content Evaluator
Evaluated and optimized large language model (LLM) outputs through high-quality RLHF (Reinforcement Learning from Human Feedback) tasks. Conducted prompt engineering evaluations, adversarial testing (red-teaming), and comparative analysis of AI-generated responses to ensure accuracy, safety, and alignment with strict project guidelines. Contributed to model training data quality assurance under a remote project management structure.