AI Trainer
In this role, I developed and evaluated prompts to assess the performance of large language models. I analyzed model outputs for accuracy, clarity, and reasoning across various specialized domains. I improved AI training datasets by identifying errors and inconsistencies in generated outputs. • Developed detailed prompt sets to measure LLM abilities across tasks. • Compared language model responses to identify strengths and weaknesses. • Provided comprehensive feedback for dataset improvement. • Focused on evaluation and error analysis for LLM outputs.