Freelance AI Data Annotator & Content Evaluator (Outlier AI)
Provided large-scale annotation and evaluation of AI-generated text by rating responses for accuracy, helpfulness, coherence, correct labeling, and instruction-following quality. Authored and refined prompts for LLMs to ensure diversity, clarity, and alignment with project guidelines. Performed side-by-side response comparisons to support RLHF pipelines and delivered structured feedback with detailed rationales. • Rated response quality across multiple evaluation dimensions • Wrote and iterated prompts for LLM training tasks • Compared response pairs for RLHF-style preference judgments • Delivered structured rationales and guidance for model improvement