Independent AI Research and Evaluation (AI Evaluator/Research Assistant style)
Performed independent evaluation of AI-generated responses using multiple LLMs to judge quality and reliability. Used detailed instructions to assess accuracy, clarity, relevance, and completeness before considering outputs usable for research or content generation. Conducted online research and fact-checking to verify claims and detect inconsistencies in AI responses. • Compared outputs from ChatGPT, Gemini, Claude, DeepSeek, and Grok • Assessed responses against accuracy, clarity, relevance, and completeness criteria • Used prompt iteration to improve output quality • Verified information via internet research and fact-checking