LLM Response Evaluation & AI Training
Performed AI training and data labeling tasks focused on evaluating and ranking LLM-generated responses. Reviewed instruction following, factual accuracy, formatting, tone, writing quality, and overall response usefulness across multilingual datasets.