AI Code Evaluator & Technical Annotator (Freelance / Remote)
As an AI Code Evaluator & Technical Annotator, evaluated and ranked AI-generated code snippets based on accuracy, efficiency, readability, and security. Conducted technical debugging, logic verification, and wrote detailed Chain-of-Thought explanations. Performed Red Teaming tests, adhered to strict annotation guidelines, and maintained accuracy in evaluation. • Evaluated Python, Java, and C++ code outputs from LLMs • Authored "Golden Responses" and step-by-step logic explanations • Performed adversarial Red Teaming on AI programming assistants • Maintained compliance with complex English instructions