Freelance Data Labeler & AI Trainer at Open Train AI
Provided Reinforcement Learning from Human Feedback (RLHF) support to improve accuracy and safety of large language models (LLMs) by refining training signals. Conducted high-precision labeling for multiple data formats according to strict, project-specific guidelines. Collaborated with global teams to identify and flag edge cases and maintain dataset robustness for AI training. • Used RLHF processes as part of LLM training workflows. • Followed detailed labeling guidelines to ensure consistent outputs. • Performed continuous quality assurance through review and refinement of annotations. • Participated in live-proctored vetting and continuous skill assessments.