AI Evaluation Specialist (Independent Contractor)
Evaluated AI-generated outputs for accuracy and consistency as part of an AI evaluation workflow. Assessed whether outputs met expected criteria and documented results to support benchmarking and reproducibility. Communicated findings through technical methodology documentation and clear reporting. • Reviewed generated text outputs against defined quality standards • Supported dataset review and benchmarking by identifying issues and gaps • Wrote reproducible methodology and findings documentation • Ensured evaluation consistency across iterations and checks