Data Labeler
Focused on testing and evaluating AI model performance across various subject areas depending on project assignments. Responsibilities included assessing the accuracy, coherence, reasoning, and instruction-following capabilities of AI-generated responses, identifying factual inconsistencies or harmful outputs, and providing detailed quality evaluations. The project contributed to improving AI model reliability and performance through human feedback, prompt testing, and annotation tasks across diverse topics and use cases.