AI Data Evaluation / Technical Reviewer (AI output and prompt evaluation)
Evaluated AI-generated code and technical proposals to determine whether they work in practice and match the existing system architecture. Assessed reasoning quality, factual/technical accuracy, and output completeness against expected production behavior. Compared model outputs to identify hidden assumptions, edge cases, and incomplete implementations. • Reviewed AI-generated code for correctness and maintainability • Evaluated prompt quality through reasoning quality checks • Compared model responses and validated technical accuracy • Produced clear, practical human-readable technical feedback