AI Trainer & Data Annotator | Freelance (Various Platforms)
Freelance AI Trainer and Data Annotator performing RLHF-based evaluation using high-quality human feedback for large language model training. Work includes writing, ranking, and refining AI-generated responses across domains such as finance, technology, writing, and reasoning tasks. Responsibilities also cover prompt engineering test design to assess model behavior, safety, and instruction-following accuracy. • Provide structured human feedback aligned to RLHF workflows. • Rank and iteratively refine model outputs for quality and alignment. • Conduct prompt engineering experiments for safety and instruction adherence. • Produce feedback reports to AI companies and maintain annotation QA standards.