/2 Stanford shows Claude and GPT generate visual reasoning without image input Stanford researchers release MIRAGE, a method to test how multimodal models use images. They take GPT-5.1, Gemini-3-Pro, Claude Opus 4.5, and Gemini-2.5-Pro and remove every image from six benchmarks.
Stanford researchers evaluate visual reasoning in multimodal models
By
–
