LLM training data support (RLHF ranking, red teaming, evaluation)
Contributed to RLHF-style dataset creation and refinement by supporting preference ranking and prompt/response feedback workflows. Participated in model red teaming activities by generating and validating ranked outputs and evaluating responses for quality and safety. Applied data quality assurance to maintain reliable training signals. • Preference ranking data preparation (RLHF) • Prompt engineering feedback and response rating • LLM evaluation and red teaming support • Data quality assurance and auditing