This paper is about how AI vision models don't have human level performance, but also this: GPT-4 without any vision capabilities at all still does remarkably well just faking it, suggesting that, much like with words, there is more regularity in the visual world then we realize
AI Vision Models: Beyond Human Performance and Visual Pattern Recognition
By
–
