For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
李伟

李伟

Model Fine-Tuning & RLHF Contributor

China flagGuangzhou, China

Key Skills

Software

No software listed

Top Subject Matter

Large Language Models
AI Agents
Quantitative Trading

Top Data Types

TextText

Top Task Types

RLHFRLHF

Freelancer Overview

Model Fine-Tuning & RLHF Contributor. Brings 4+ years of professional experience across complex professional workflows, research, and quality-focused execution. Core strengths include Internal and Proprietary Tooling. Education includes Bachelor of Science, 中山大学 (2016). AI-training focus includes data types such as Text and labeling workflows including RLHF.

Labeling Experience

Model Fine-Tuning & RLHF Contributor

TextTextRLHFRLHF

Participated in model fine-tuning and reinforcement learning from human feedback (RLHF) processes for large language model training. Collaborated in the evaluation, rating, and optimization of model outputs to improve language generation quality. Contributed to iterative model improvement cycles focusing on AI agent systems and quant trading modules. • Directly involved in RLHF as part of end-to-end model training. • Improved output quality through systematic annotation and human feedback. • Used internal/proprietary tools alongside PyTorch and TensorFlow. • Targeted applications include AI agents and quantitative trading AI solutions.

2023 - Present

Education

中山大学

Bachelor of Science, Computer Science and Technology

Bachelor of Science
2012 - 2016

Work History

D

DeepSeek

AI Development Engineer

Guangzhou
2023 - Present