Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Industry & Money Story
Industry & Money

AI models confidently describe images they never saw, and benchmarks fail to catch it

Via The Decoder
Monday, Mar 30, 2026 · 3:53PM
Summary

Multimodal AI models like GPT-5, Gemini 3 Pro, and Claude Opus 4.5 generate detailed image descriptions and medical diagnoses even when no image is provided. A Stanford study shows that common benchmarks obscure the problem. The article AI models confidently describe images they never saw, and bench

Continue reading the full article
Read at The Decoder
the-decoder.com
Back to all stories