Senior AI Content Evaluator
As a Senior AI Content Evaluator at Scale AI, I evaluated and ranked LLM-generated text for tone, clarity, fluency, factual accuracy, and suitability for various audiences. My work included writing detailed preference feedback and justifications integral to fine-tuning models via RLHF pipelines. I designed and crafted prompts for creative, conversational, persuasive, and knowledge-based scenarios for large-scale AI training. • Evaluated subtle language issues in LLM outputs for clarity and appropriateness. • Consistently achieved inter-annotator agreement scores above 97%. • Collaborated with QA leads to refine annotation guidelines and editorial style guides. • Applied rigorous standards to maintain team quality benchmarks and feedback processes.