Prompt Evaluation Project
Reviewed and rated AI responses based on accuracy, tone, and relevance as part of a prompt evaluation workflow. Helped improve conversational AI performance by applying consistent evaluation criteria and providing structured feedback. Ensured evaluation outputs aligned with quality expectations for downstream model improvement. • Assessed response accuracy, tone, and relevance for given prompts. • Applied rubric-based judgment to identify improvements in outputs. • Provided feedback to support iterative improvements to conversational AI. • Ensured consistent rating quality across tasks.