Principal Software Engineer, Netsmart Technologies
Architected a human-in-the-loop data validation platform that processed millions of records per month and routed expert evaluation tasks by domain, language, and prior quality history. Integrated LLMs into clinical documentation workflows by creating standardized prompt templates, output scoring rubrics, and inter-rater reliability dashboards for consistent evaluation. Led development of APIs and evaluation tooling that supported large-scale expert validation and quality measurement for AI outputs. • Human expert task routing and data validation • Creation of prompt templates and scoring rubrics • Inter-rater reliability dashboards for evaluator consistency • Large-scale processing of labeled/evaluated clinical records