RLHF Reviewer / SFT Auditor (March 2026 - Present)
Evaluated LLM responses by applying precise data labeling to improve output quality and reliability. Used RLHF-based ranking to assess helpfulness, truthfulness, safety, and adherence to constraints. Authored reference responses and golden datasets to support supervised fine-tuning and advanced domain reasoning. • Ranked model outputs using RLHF methodologies for evaluation and grading • Performed fact-checking and hallucination detection to validate responses • Conducted quality assurance checks on tone, accuracy, and logical consistency • Produced labeled reference materials (golden datasets) for training support