Daily AI Tester (Personal Practice)
Tested text and image generation behavior by experimenting with different prompt formulations and instruction constraints. Tracked failure cases such as complete disregard of rules (e.g., not including people) and mistakes inside generated images. Compared instruction-following, quality, and safety differences between ChatGPT and Gemini over time.• Ran weekly experiments to probe model limits.• Documented when models ignored prohibitions or created errors.• Assessed how lighting, camera angles, and objects were handled.• Monitored quality and safety rule behavior across two AI tools.