AI Agent Trainer & Evaluation Specialist
Participated in the training loop for financial and audit-focused AI agents through expert-level human feedback. Evaluated the accuracy, consistency, and compliance of agent outputs in complex computational and decision-making scenarios. Identified and corrected model hallucinations and logical inconsistencies in high-stakes business tasks. • Evaluated multi-step agent workflows for correctness and tool usage in automated tasks. • Audited agent decision-making for IT controls and regulatory compliance. • Provided detailed human feedback to refine model behavior and output quality. • Contributed to model alignment by diagnosing and solving edge cases and adversarial challenges.