Self-study: RLHF / LLM human feedback training and dialogue evaluation labeling practice
Self-studied RLHF and LLM human feedback training with a strong focus on dialogue evaluation and data labeling standards. Performed 500+ LLM dialogue evaluation and labeling practice covering quality scoring, ranking/sorting, and content safety checking. • Evaluated and scored dialogue responses based on guidelines • Sorted and ranked outputs using preference-style judgments • Conducted safety review and content QA • Practiced applying RLHF-style feedback and labeling criteria