Senior AI Content Evaluator (Freelance/Contract)
Performed RLHF evaluation and generation tasks including response ranking and rewriting for LLM development platforms. Audited AI-generated code and technical explanations to catch logic errors and security vulnerabilities while ensuring compliance with multi-layer annotation standards. Supplied “Golden Response” examples to guide consistent model training data quality across the taxonomy. • Conducted high-level preference/ranking judgments for model outputs. • Performed rewriting based on guideline adherence and quality thresholds. • Performed QA-style error and risk checks on generated technical content. • Produced reference outputs (“Golden Responses”) for training calibration.