AI Data Trainer & Content Specialist (LLM data labeling, evaluation, and response auditing) at CrowdGen (Formerly Appen)
Reviewed and annotated large-scale query-response training datasets to support LLM evaluation and search relevance improvements. Performed quality assurance by scoring and correcting model outputs to maintain accuracy above 95% across multiple projects. Worked with prompt testing and semantic analysis workflows to ensure responses met guideline requirements. • Evaluated 15,000+ complex search query-response pairs and training data points • Tested prompts, audited AI responses, and corrected inaccurate or low-quality outputs • Conducted search relevance evaluation, fact-checking, and semantic analysis • Followed and applied dense project guidelines quickly to pass qualification testing