AI Reasoning Expert / Evaluator
Assessed and scored large language model outputs for expert-level mathematical reasoning tasks using rubric-based evaluation standards. Provided structured feedback focused on correctness, clarity, and reasoning integrity, including error categorization and justification writing. Worked on multimodal reasoning evaluations involving text plus image interpretation and cross-modal reasoning validation. • Trained and evaluated LLMs on multi-step mathematical reasoning • Performed rubric-based scoring with written justifications • Detected, categorized, and corrected reasoning errors in solutions • Supported multimodal projects (text, image, and reasoning) and Meta-focused initiatives.