AI Training and Evaluation Specialist (Freelance)
Designed and refined prompts to improve AI model accuracy and consistency across complex, multi-step workflows. Analyzed model outputs to identify failure patterns and edge cases that degrade performance. Built evaluation frameworks to measure reliability and guide iterative improvements in real-world usage. • Prompt design and refinement for workflow accuracy • Output analysis to surface failure patterns • Evaluation framework development for performance measurement • Cross-project collaboration to enhance reliability