Analyst-AI LLM Practice (Contract)
Worked as an Analyst in the AI LLM Practice team on a contract focused on evaluating and improving Large Language Models (LLMs). The work included analyzing model outputs for accuracy, relevance, safety, and guideline compliance while performing validation and quality checks. I followed detailed annotation guidelines and supplied structured feedback to support model improvement and quality benchmarks. • Evaluated AI-generated responses for accuracy and relevance. • Assessed responses for safety and compliance with project guidelines. • Performed data validation and content classification to improve performance. • Provided structured annotation feedback for iterative LLM enhancement.