AI Content Annotator & Prompt Evaluator (Freelance)
Freelance AI Trainer and data labeler performing LLM output evaluation and annotation for quality and alignment-related criteria. Responsibilities included rating and comparing AI-generated text responses using rubrics spanning helpfulness, accuracy, and safety. The work also supported prompt-based workflows by iteratively testing prompts and evaluating results across multiple iterations. • Rated and compared LLM responses for quality dimensions (helpfulness, accuracy, safety) • Performed RLHF-style evaluation to support model improvement • Wrote, tested, and iteratively refined prompts for better outputs • Applied consistent guideline-following during annotation at scale