高级AI研发工程师(LLM 数据评估与RLHF数据管理)
This role focused on evaluating and optimizing large language models (LLMs) for various AI applications. I conducted data quality control, performance assessment, and contributed to RLHF data collection and model safety evaluation. My daily work involved prompt engineering, code review, and chain-of-thought experiments. • Performed LLM evaluation for code generation and mathematical reasoning tasks. • Led RLHF data collection, annotation, and quality assurance. • Designed and executed model safety and red teaming workflows. • Processed and validated large-scale multi-modal data for AI development.