AfterQuery Expert (Data Annotator)
Validated AI trajectory outputs through multi-stage rubric-based checks for determinism, alignment, and technical accuracy. Identified and mitigated ambiguity signals by tightening constraints and refining gold-standard answers. Curated and standardized professional benchmark artifacts and authored SWE-bench-style coding tasks for model capability evaluation. • Graded complex rubric outputs to ensure factual and technical correctness. • Reduced ambiguity by reframing questions and updating high-value reference answers. • Standardized benchmark materials (DOCX, PPTX, XLSX) for consistent evaluation. • Built automated test harnesses to verify generated code against reference implementations.