AI Trainer at Aidralabs (remote) specializing in data annotation/labeling, AI feedback, and RLHF-based evaluation
Trained and evaluated AI/ML models by producing high-quality labeled data and structured human feedback. Delivered prompt engineering work and participated in RLHF sessions involving rating, ranking, and correcting AI-generated outputs. Supported continuous model improvement through quality assurance collaboration and targeted evaluations such as bias detection and factual accuracy reviews. • Produced and validated labeled training data for AI model use • Wrote and refined prompts to improve response quality and accuracy • Performed RLHF feedback: rating, ranking, and correction of outputs • Coordinated with QA teams to ensure consistency and accuracy in datasets