AI Trainer Expert (Micro1) — Trees Teak, Realm Isbilia RLE, Realm Isbilia STEM
Performed end-to-end AI evaluation by reviewing prompts, supporting documents, workflow instructions, and applying grading rubrics to benchmark AI performance on real-world computer-based tasks. Analyzed AI agent outputs and behaviors, producing structured observations and high-quality evaluation reports based on reviewer feedback. Refined prompts, rubrics, and grading criteria with the team to improve task quality and benchmarking reliability across evolving domains and project requirements. • Worked with prompt and rubric driven evaluation workflows for AI tasks. • Produced concise performance summaries and structured observation notes. • Updated assessment standards by incorporating reviewer and team collaboration feedback. • Adapted evaluation frameworks across multiple domains to stay aligned with requirements and quality criteria.