Freelance AI Training Specialist & Data Annotator (Remote Contracts)
Trained and evaluated large language model (LLM) responses using Reinforcement Learning from Human Feedback (RLHF) methodologies. Reviewed AI outputs for factual precision, logical syntax structure, grammatical coherence, and compliance with specified technical constraints. Authored prompt constraints and structured datasets to guide conversational agents toward professional, error-free customer interactions. • RLHF-based response evaluation and optimization • Quality checks for factual/logical/grammatical compliance • Prompt-constraint authoring for dialogue behavior • Dataset creation for training conversational agents