RLHF Specialist for Large Language Models
Contributed to the fine-tuning of a Large Language Model by ranking and rewriting AI-generated responses. I evaluated outputs based on Criteria, e.g., truthfulness and safety] and provided detailed reasoning for each ranking. I completed over 5000 tasks, adhering to strict style guides to ensure a natural and helpful tone.
2024 - 2026