AI Quality Evaluator - DataAnnotation Remote
You evaluate AI-generated responses for correctness, reasoning quality, and factual accuracy across a variety of tasks. You validate code, mathematical solutions, and natural-language outputs to support quality assurance. You apply prompt-engineering techniques to improve large language model performance and ensure reliable results in production-style workflows. • Review and assess AI outputs for factual accuracy and logical consistency • Check structured answers including code and math solutions • Use prompt-engineering methods to refine model behavior • Contribute to QA workflows for large language models