AI-generated Code Reviewer and LLM Evaluation Lead
Led comprehensive AI-generated code review for LLM training datasets, focusing on quality, correctness, and best practices. Developed and utilized an internal LLM evaluation framework to assess model output alignment and effectiveness. Oversaw annotation workflow management and quality assurance for code-based data labeling. • Evaluated AI-generated Python and JavaScript code for accuracy and adherence to guidelines. • Managed dashboards for ML annotation workflow and QA tracking. • Contributed to model alignment efforts for improved LLM outputs. • Provided subject matter expertise in code reviewing for AI training.