For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
D

Dx L.

Self-study: RLHF / LLM human feedback training and dialogue evaluation labeling practice

China flagChina, China

Key Skills

Software

No software listed

Top Subject Matter

LLM human feedback
Rlhf Domain Expertise
dialogue evaluation

Top Data Types

TextText

Top Task Types

RLHFRLHF

Freelancer Overview

Self-study: RLHF / LLM human feedback training and dialogue evaluation labeling practice. Core strengths include N and A. AI-training focus includes data types such as Text and labeling workflows including RLHF.

Labeling Experience

Self-study: RLHF / LLM human feedback training and dialogue evaluation labeling practice

TextTextRLHFRLHF

Self-studied RLHF and LLM human feedback training with a strong focus on dialogue evaluation and data labeling standards. Performed 500+ LLM dialogue evaluation and labeling practice covering quality scoring, ranking/sorting, and content safety checking. • Evaluated and scored dialogue responses based on guidelines • Sorted and ranked outputs using preference-style judgments • Conducted safety review and content QA • Practiced applying RLHF-style feedback and labeling criteria

2025 - Present