AI Trainer & AI Evaluation Contributor (AI evaluation of LLM outputs)
Evaluated AI-generated outputs for accuracy, reasoning quality, factual consistency, clarity, and instruction-following across multiple knowledge domains. Reviewed prompts and corresponding model responses with a focus on business, research, customer service, operational, and general knowledge suitability. Contributed to an AI evaluation workflow that emphasizes quality assessment of language model outputs.• Assessed response correctness and reasoning quality.• Checked for factual consistency and clarity.• Verified adherence to instructions and prompt intent.• Performed domain-spanning prompt/response reviews.