Agentic AI Workflow Evaluator at Turing (Remote)
Executed computer-based agent workflow evaluations by recording full task sessions and documenting click-by-click event traces. Assessed task efficiency, completeness, and adherence to expected workflow standards while identifying breakdown points in multi-step agent execution. Applied structured annotation protocols and provided actionable reliability feedback to improve workflow optimization.• Recorded complete task sessions with interaction traces and event documentation.• Evaluated agent workflow efficiency, completeness, and compliance to standards.• Flagged UI/UX inconsistencies and reasoning gaps in automated tasks.• Used structured annotation for reproducibility and benchmarking alignment.