Ranking & Evaluation (AI responses quality rating) at Invisible technologies
Provided response ranking for two or more AI outputs, evaluating truthfulness, safety, and helpfulness. • Compared prompt-response pairs and selected the most accurate, safe, and helpful answer • Performed evaluation based on quality criteria such as truthfulness and safety • Supported iterative improvements to AI behavior through consistent ratings • Helped streamline AI response assessment workflows across projects