Technical AI Training Specialist (Freelance) — Outlier/Scale AI
Technical Data Specialist and RLHF Contributor within the Scale AI/Outlier ecosystem, specializing in SFT and Reinforcement Learning from Human Feedback for Tier-3 coding and reasoning models. Developed high-signal datasets through complex prompt engineering, multi-turn adversarial testing, and ground-truth generation. Consistently maintained a 4.5/5.0 quality rating while performing rigorous code debugging (Python) and logical reasoning evaluations. Executed $100+ in technical tasking, ensuring strict adherence to HHH (Helpfulness, Honesty, Harmlessness) safety alignment and factuality protocols for major LLM deployments.