Freelance Tech Support & Data Projects (AI output evaluation and tutoring)
Evaluated AI-generated technical outputs for student users by checking code correctness and identifying hallucinations. Performed quality review using structured judgment to assess whether responses matched expected logic and requirements. Documented observed failure modes and used findings to improve prompt and evaluation approaches for future outputs. • Reviewed AI tool outputs for 30+ student users. • Flagged hallucinations and incorrect code in about 40% of test cases. • Wrote step-by-step Python and SQL explanations to support debugging and learning. • Maintained timely delivery while managing multiple clients with varying requirements.