AI/Speech evaluation and rating (prompt testing and QA) based on strict guidelines
Evaluated and rated AI-generated text and audio responses using strict project guidelines to judge accuracy, relevance, and natural language flow. Provided objective, constructive feedback on AI interactions to support continuous improvements of machine learning behavior and outputs. Performed quality and safety-focused review by checking for potential logic issues and ensuring responses meet stated criteria. • Rated AI responses against rubric-style criteria for correctness and contextual alignment • Analyzed contextual relevance and natural language quality for written outputs • Assessed audio and speech responses for quality and conformity to requirements • Identified potential safety or logic flaws through prompt-based testing and review