AI Model Evaluator & Software Data Rater (Turing)
Trained and evaluated advanced AI model responses through structured testing focused on clarity, accuracy, safety, and helpfulness. Reviewed and labeled complex datasets within remote evaluation workflows to optimize model behavior for real-world deployment. Applied consistent labeling criteria aligned to safety and usefulness requirements.• Rated responses for helpfulness, clarity, and correctness.• Assessed safety constraints and compliance of outputs.• Labeled complex dataset examples for model improvement.• Collaborated in remote workflows to meet evaluation standards.