AI Data Quality Specialist
As an AI Data Quality Specialist, I evaluated large volumes of LLM-generated text for accuracy, coherence, and alignment with human values. I contributed to RLHF prompt dataset creation and developed annotation guidelines to improve consistency among annotators. My work involved curating safety datasets, identifying problematic content, and mentoring junior team members. • Evaluated over 500 LLM outputs weekly using rubric-based assessments. • Created prompt–response pairs for RLHF data in STEM, law, and creative writing. • Improved inter-annotator agreement by 18% via refined guidelines. • Documented model hallucinations and edge cases for safety data curation.