AI Trainer & Content Specialist — Data Annotator ([email protected])
Performed RLHF/SFT data creation, prompt engineering, and response evaluation as an AI trainer and content specialist/data annotator. Built instruction-following and style-guide compliant prompt-response pairs intended for training and preference modeling. Conducted response ranking and QA/data validation to improve output quality and reliability. • Create RLHF/SFT training examples and preference-oriented materials • Engineer prompts and evaluate/rank responses for quality • Ensure instruction-following accuracy and strong style adherence • Validate outputs using QA/data validation practices