Independent AI Projects Specialist (Data Training & Annotation)
As an independent specialist, I led continuous improvement projects aimed at optimizing large language models through high-quality data curation and reinforcement learning from human feedback. My work involved semantic annotation, evaluation of model outputs, and detailed bias/error detection to enhance model alignment and safety. I developed prompt engineering strategies and structured datasets for training and fine-tuning purposes. • Evaluated and classified LLM outputs for logical consistency and user alignment. • Engineered and refined prompts for complex reasoning and generation tasks. • Identified and documented hallucinatory or biased model outputs for feedback loops. • Structured, enriched, and categorized text datasets for supervised learning and RLHF cycles.