AI Tool Evaluation & Prompt Testing (project)
Evaluated AI-generated outputs and conducted prompt testing to improve model response performance. Compared results across multiple AI systems and adjusted prompts to strengthen clarity, relevance, and usefulness. Focused the evaluation on measurable response quality attributes such as accuracy and usefulness. • Tested prompt variations and assessed the impact on response quality • Compared outputs across multiple AI systems/models for consistency • Rated responses on relevance, clarity, and usefulness criteria • Iterated prompts based on observed shortcomings and error patterns