AI Dynamics

Global AI News Aggregator

About

Stanford researchers evaluate visual reasoning in multimodal models

/2 Stanford shows Claude and GPT generate visual reasoning without image input Stanford researchers release MIRAGE, a method to test how multimodal models use images. They take GPT-5.1, Gemini-3-Pro, Claude Opus 4.5, and Gemini-2.5-Pro and remove every image from six benchmarks.

→ View original post on X — @alphasignalai