AI Model Evaluation Specialist (Contractor) - Handshake AI
Conducts AI evaluation and training work centered on prompt adherence, output quality, reasoning, formatting, and instruction-following. Reviews model outputs to identify hallucinations, weak reasoning, missing constraints, and inconsistent formatting. Performs high-volume grading and refinement tasks while learning current model behavior and improving evaluation workflows. • Evaluate outputs for quality, reasoning, formatting, and failure modes • Provide structured feedback to refine prompts and system behavior • Detect hallucinations, constraint gaps, and inconsistencies • Execute repeatable, high-throughput evaluation workflows