AI/Data Labeling & Annotation (Education Project) — RLHF & RLAIF, Model Evaluation
This experience focuses on AI optimization activities related to RLHF and RLAIF for model improvement. The work includes evaluating model behavior and participating in red teaming to identify weaknesses and areas for refinement. It supports iterative alignment and performance tuning of AI systems for more reliable outputs. • RLHF & RLAIF model optimization • Model evaluation for quality and safety • Red teaming to surface failure modes • Ongoing AI training and alignment work