AI Trainer & Evaluator | Handshake AI Fellowship
Developed and evaluated domain-specific prompts to assess large language model performance in specialized areas. Reviewed model outputs for scientific accuracy, clarity, and depth to guide improvements. Provided expert feedback to enhance AI understanding for downstream usage. • Prompt design and iteration for domain coverage • Qualitative evaluation of LLM responses • Expert review and constructive feedback • Focus on accuracy, clarity, and depth