AI Generalist (LLM response evaluation and dataset training support)
Contributed to AI training by evaluating and refining LLM-related prompt and response content as part of dataset preparation. Conducted response evaluation, ranking/comparison, and guideline-based quality checks to improve dataset reliability. Performed careful attention to detail and edge-case handling to ensure high-quality training examples. • Evaluated prompts and responses for quality using established rubrics • Ranked and compared outputs to identify best candidate responses • Performed edge-case analysis to handle ambiguous or tricky examples • Ensured compliance with strict guidelines during annotation/evaluation