Freelance AI Data Specialist & Evaluator (Mindrift RLHF, AtlasCapture video annotation, Micro1 prompt specialist)
Conducted reinforcement learning from human feedback (RLHF) tasks including editing model responses for logic, factuality, and style adherence as part of human-in-the-loop refinement. Completed video data annotation by identifying temporal actions and ensuring precise timestamping for computer vision training. Delivered prompt stress-testing support using complex chain-of-thought prompt strategies to evaluate LLM reasoning robustness. • Edited and validated agent responses for logic, factuality, and style compliance. • Performed RLHF human review to refine agent behaviors. • Annotated video segments with action timing and accurate timestamps. • Designed complex prompt chains to stress-test LLM reasoning capabilities.