AI model interaction and evaluation (ongoing)
Actively uses and critically evaluates outputs from advanced AI models such as ChatGPT and Grok to assess accuracy, coherence, and reasoning quality. Refines prompts and iteratively improves AI-generated content while maintaining defined logical and quality standards. Provides human feedback loops to support AI training and evaluation through reasoning-oriented assessment tasks. • Critically evaluate AI outputs for quality • Refine prompts to improve responses • Maintain logical consistency and quality standards • Support AI training via human feedback and reasoning evaluation