AI Response Evaluator & RLHF Specialist
Evaluated LLM-generated responses based on accuracy, relevance, clarity, tone, and factual consistency using RLHF principles. Performed detailed content quality reviews and data annotation across dual-language workflows. Compared multiple AI outputs to select the most helpful, complete, and reliable responses while meticulously identifying logical inconsistencies, factual errors, and strict guideline violations. Provided structured, objective feedback to improve model alignment and instruction-following capabilities.