AI Evaluator & Bilingual Prompt Engineer
As an AI Evaluator and Bilingual Prompt Engineer at DataAnnotation.tech, I evaluated and aligned Large Language Models (LLMs) using human feedback. My work involved designing complex prompts in both Arabic and English, as well as conducting rigorous A/B testing on AI-generated responses. I provided comprehensive rationale justifications for each evaluation, directly impacting model fine-tuning. • Designed and tested prompts for safe and accurate LLM behavior. • Performed detailed factual and logical assessments of AI responses. • Drafted structured evaluation reports for model alignment. • Specialized in bilingual prompt engineering and RLHF methodologies.