Senior Full Stack Engineer & AI Architect — Contract (Private) (01/2026 – Present)
Validated Claude outputs in a Python workflow using structured feedback loops to improve correctness and reliability. Built Python evaluation pipelines to assess AI-generated code for correctness, efficiency, and security. Produced evaluation reports that identify architectural and performance risks and guide remediation work with AI engineering teams. • Performance, correctness, and security-focused evaluation of model outputs • Structured feedback loop design to improve reliability • Generation of risk and remediation reports for engineering alignment • Ongoing evaluation pipeline development in Python