Data Quality Auditor & AI Evaluation Reviewer (Independent Contractor / Remote)
Performed multi-pass human-in-the-loop evaluation of complex multi-turn AI reasoning trajectories to detect logic loops, factual deviations, and output truncation across long step sequences. Scored model behaviors using enterprise rubrics covering factual correctness, algorithmic efficiency, and logical consistency to support quality assurance and iteration. Verified grounding and safety compliance by auditing context data pipelines to prevent future-dated fabrication and guideline leakage while flagging hallucinated syntax and model overreach.• Evaluated 20+ step trajectories for reasoning quality and stability• Applied rubric-based scoring and classification to annotate model behavior• Conducted compliance checks for policy, safety, and grounding adherence• Identified and reported model logic errors for remediation