Independent AI Language Trainer & Evaluator
As an independent AI Language Trainer & Evaluator, I assessed and ranked LLM-generated responses for accuracy, tone, and logic. I engineered prompts to test AI behavior across various linguistic and cultural contexts and reviewed AI-generated code for correctness. I also categorized datasets for bias reduction and safe model alignment. • Conducted RLHF-based evaluations on Arabic and English LLM outputs. • Designed complex prompt scenarios to evaluate edge cases. • Performed technical reviews of backend code and database logic. • Applied formatting and QA protocols to align labeled data with project standards.