AI prompt Testing (ClaudeAI)
Tested AI behavior by running prompt variations and analyzing responses for reasoning quality and accuracy. Identified errors and hallucinations in model outputs and proposed improvements to make answers clearer. Collaborated with a team to address product/service issues using evaluation findings. • Designed and tested prompts to evaluate AI reasoning and accuracy • Analyzed responses to detect errors and hallucinations • Suggested prompt/answer improvements for clarity and correctness • Coordinated with teammates to improve performance and outcomes