AI Prompt Evaluation & Response Quality Assessment
Conducted prompt engineering, AI response evaluation, and quality assurance testing for a conversational AI application. Assessed model outputs for instruction adherence, factual accuracy, reasoning quality, language fluency, formatting consistency, and user intent alignment. Identified and documented errors, including hallucinations, logical inconsistencies, incomplete responses, and policy compliance issues. Provided structured feedback and annotations to support continuous improvement of AI-generated outputs and user experience.