AI Trainer & Data Quality Specialist
I contributed to the optimization of a Large Language Model (LLM) through high-quality Supervised Fine-Tuning (SFT) and RLHF workflows. My work involved generating complex instruction-following datasets and evaluating multi-turn dialogues for accuracy, safety, and helpfulness. I specialized in Chain-of-Thought (CoT) reasoning tasks to improve the model's logical consistency in technical domains. During this period, I maintained a consistent 95%+ quality audit score, successfully identifying and correcting model hallucinations across thousands of prompts.