AI Response Evaluation and Data Annotation for Large Language Models
Worked on AI training and data annotation projects involving the evaluation and improvement of large language model outputs. Responsibilities included reviewing AI-generated responses for factual accuracy, relevance, clarity, instruction-following, and overall quality. Annotated and categorized text data according to detailed project guidelines, identified inconsistencies, and provided feedback to improve model performance. Performed tasks such as prompt-response evaluation, content classification, ranking model outputs, quality assurance reviews, and data validation. Maintained high annotation accuracy while meeting project deadlines and adhering to strict quality standards. Contributed to the development of safer, more reliable, and more helpful AI systems through consistent and accurate human feedback.