高级人工智能训练师(GLM系列RLHF数据标注与质检)
Responsible for GLM series LLM RLHF human feedback data collection and quality control, producing over 50,000 high-quality annotated entries. Designed and optimized prompt engineering strategies to improve accuracy on Chinese reasoning tasks. Led a small annotation team to construct multi-turn dialogue datasets with strong consistency and shorter delivery cycles. • GLM RLHF data collection and quality assurance • Prompt engineering for Chinese reasoning accuracy improvement • Multi-round dialogue dataset construction with consistency targets • Participation in internal annotation standards and QC process formulation