AI Data & Evaluation Specialist - Turing
Review, grade, and refine AI model outputs for enterprise clients, validating technical logic, factual accuracy, and code-generation samples across multiple frameworks. Identify, document, and patch edge-case logical failures and model hallucinations by writing definitive "golden responses" to re-train neural networks. Deconstruct multi-page annotation style guides to ensure 100% compliance across massive data-labeling batches, achieving a consistent 99.4% quality assurance rating. Author clear, highly structured markdown justifications and source citations explaining why specific model responses were penalized or approved.