AI Training & Data Labeling Specialist — Handshake AI (Remote)
Provided evaluation and training feedback for AI responses in support of reinforcement learning workflows. Created and refined labeled datasets for RL training by assessing accuracy, reasoning quality, and contextual relevance. Flagged hallucinations and logical inconsistencies and produced written feedback to improve model performance and reliability. • Evaluated AI-generated responses for accuracy, reasoning quality, and contextual relevance • Created and refined training datasets for reinforcement learning workflows • Labeled and classified structured and unstructured business data for AI improvement • Identified hallucinations, incomplete outputs, and logical inconsistencies