AI Prompt & Evaluation Specialist – Freelance (Various Projects, Remote)
Freelance AI Prompt & Evaluation Specialist work focused on developing and evaluating LLM training materials and response quality. They applied defined scoring criteria to assess model outputs and identified flawed responses requiring correction. The work contributed to improved factual accuracy and explainability for coding-related AI tasks. • Developed training data and prompt sets for LLMs using OpenAI API and Anthropic. • Created technical prompts to test AI reasoning and code generation. • Scored and evaluated model responses against quality criteria. • Annotated incorrect or insufficient outputs with corrected examples.