AI Safety and Evaluation Analyst (Contract) - Dataannotation
Evaluated large language model outputs for accuracy, safety, and reliability while identifying and documenting key failure modes. Conducted targeted research and fact verification to reduce misinformation risk in high-volume evaluation workflows. Applied structured evaluation frameworks and judgment-based criteria to assess ambiguous and edge-case scenarios and support consistent safety alignment. • Reviewed model responses for hallucinations, prompt-injection vulnerabilities, and unsafe content • Developed adversarial prompt strategies to expose weaknesses in model behavior • Documented evaluation decisions and escalated high-risk findings for improved model performance