AI Data Annotator - Handshake AI
Evaluated how well language models reason, write, and follow instructions. Rated and ranked content against defined relevancy and quality scales, applied strict guidelines to judge whether outputs met criteria, and flagged ambiguous items for review — sustaining accuracy and throughput across large queues using rubric-based scoring and calibration against reference examples.