Lead Software Engineer (AI-assisted coding evaluation and model output validation) — Arcadia Byte Labs
Led development of AI-assisted code evaluation tooling to validate and debug AI-generated programming outputs. Assessed model-generated reasoning by analyzing recurring failure modes and documenting logical consistency checks. Standardized reproducible debugging workflows to improve reliability and engineering quality across distributed systems. • Performed evaluation of AI-generated code outputs for correctness and consistency • Analyzed and recorded recurring reasoning failures with AI research teams • Built automated validation and debugging frameworks for defect reduction • Documented reproducible workflows for repeatable model evaluation and debugging