AI Training Specialist — Alignerr (2025–Present)
Performed LLM output evaluation to assess quality, reasoning, and consistency against desired criteria. Optimized training datasets by analyzing model behavior and updating example selections and associated annotations. Helped improve training workflow reliability by identifying failure patterns such as low-quality responses and hallucinations. • Evaluated LLM responses for correctness and alignment • Improved dataset validation through iterative review • Supported model reasoning enhancements via targeted data updates • Reduced hallucination tendencies through quality-focused evaluation