AI Response Evaluator (Human Trainer)
Evaluated AI-generated responses using a human-in-the-loop RLHF style rubric and assigned quality ratings on a 1–7 scale. Assessed alignment with desired model behavior across diverse conversation types, focusing on tone, helpfulness, honesty, writing style, and instruction-following accuracy. Produced structured, segment-level feedback by identifying both aligned and misaligned parts and providing written rationales for each label. • Rated response quality on a 1–7 quality scale. • Applied rubrics covering tone, helpfulness, honesty, style, and instruction adherence. • Wrote well-reasoned rationales highlighting aligned/misaligned segments. • Composed 4–5 sentence user-intent summaries per task to reflect conversational context.