Generalist AI Trainer – Stellar AI (Apr 2025 – Present)
Delivered high-accuracy annotation, evaluation, and prompt engineering for STEM, education, and general AI use cases. Assessed LLM outputs for accuracy, coherence, and logical consistency to improve reliability, while identifying hallucinations and weak reasoning patterns. Supported RLHF workflow improvements by refining training datasets and contributing to safety testing. • Produced and refined JSON-based training/evaluation datasets for LLMs • Performed structured output validation and rubric-driven evaluation • Conducted red-teaming/safety testing to identify vulnerabilities • Collaborated with cross-functional teams to improve annotation efficiency