AI Data Specialist, Oak Ridge National Laboratory – AI Research Partnership (Remote)
Annotated and validated technical computing records to train large language models for English-language performance. Conducted quality audits of AI-generated code explanations and technical documentation against domain standards. Performed bias and factual accuracy reviews across diverse datasets to ensure reliable outputs for technical audiences. • Created custom annotation guidelines and quality rubrics. • Improved inter-annotator agreement from 78% to 94% via structured guidance. • Reduced hallucinations by 41% while achieving 96% adherence to domain standards. • Contributed structured feedback loops that supported multiple model iteration cycles.