AI Data Annotator & RLHF Specialist (Scale AI / Remotasks)
Performed RLHF preference-ranking by evaluating and ranking model-generated outputs for reinforcement learning guidance. Completed 1,000+ annotation tasks across text classification, entity recognition, and AI response ranking while maintaining quality above 95%. Reviewed and improved AI-generated code snippets by identifying logical errors, security issues, and stylistic inconsistencies. • RLHF preference ranking and output evaluation for major LLMs • Text classification and entity recognition annotations • Code review support for Python, JavaScript, and SQL • Sensitive content moderation and safety dataset labeling within strict guidelines