LLM Evaluator & Annotator
As an LLM Evaluator & Annotator at Cognito AI, I evaluate and compare AI-generated responses for multiple criteria using structured guidelines. I detect hallucinations, factual errors, and policy violations to ensure safety and quality. This work directly refines large language model (LLM) reasoning and output quality. • Assessed outputs for accuracy, safety, clarity, and helpfulness • Flagged and reported violations or inconsistencies for model alignment • Applied human judgment across diverse knowledge domains • Provided structured feedback through scoring rubrics for LLM refinement