RLHF AI Trainer (Generalist, Malayalam, Maths)
As an RLHF AI Trainer, I specialized in reinforcement learning with human feedback for large language model (LLM) training. My contributions focused on generating, validating, and evaluating prompt/response pairs in Generalist, Malayalam, and mathematical reasoning tasks. I also undertook prompt engineering and task-specific coding to improve LLM data workflows. • Conducted prompt engineering to create high-quality training data for LLMs. • Provided data validation and evaluation to refine model outputs in text-based domains. • Delivered training signals and feedback in mathematical reasoning and Malayalam NLP tasks. • Collaborated via remote freelance workflows and documented task-specific outputs.