AI Research Agent & Performance Evaluation System (autonomous research agent with rubric-based evaluation)
Built an autonomous AI agent that decomposes and completes multi-step research tasks using external tools. Developed the agent workflow to manage tool use and reasoning steps required to produce final research outputs. Emphasized evaluation-driven iteration by pairing task execution with scoring rubrics. • Implemented agentic multi-step task breakdown • Integrated external tool usage for completion • Produced final research outputs with consistency checks • Iteratively refined behavior guided by evaluation results