LLM Evaluation & Safety Certification (Anthropic via Skilljar)
You were trained by Anthropic (via Skilljar) to evaluate and rank Large Language Model outputs. The training focuses on safety, helpfulness, and factual accuracy to support responsible deployment. You applied evaluation criteria to guide consistent quality judgments for model responses. • Score/rank LLM outputs by safety • Judge helpfulness of responses • Assess factual accuracy • Use criteria aligned with LLM safety standards