AI Prompt Writer / AI Model Evaluator (Outlier AI, Freelance/Contract, Remote)
Performed AI model evaluation and fine-tuning support across generalist and legal tasks by assessing reasoning quality, accuracy, and instruction-following. Designed complex prompts to probe model weaknesses and compared outputs against structured task requirements. Reviewed and rewrote model responses to improve clarity, factual accuracy, structure, tone, and usefulness. • Developed evaluation rubrics covering reasoning quality, completeness, legal analysis, factual accuracy, and safety • Conducted response assessment to identify gaps in legal-focused issue analysis and drafting-style outputs • Collaborated with a remote global contributor network while maintaining confidentiality and meeting quality targets