RLHF (Reinforcement Learning from Human Feedback) for Advanced Data Analytics & Code Generation
Served as a domain-expert annotator to train a specialized Large Language Model (LLM) capable of generating advanced data analytics workflows and automation scripts. The role involved creating complex technical prompts, evaluating model-generated code (SQL, R, VBA, and Excel formulas) for functional accuracy, and ranking outputs based on safety, efficiency, and clarity. This high-quality human feedback directly improved the model's ability to interpret analytical requests and deliver bug-free, deployment-ready code.