From my initial testings GPT-5 is as bad as GPT-4o on instruction following. This is a real bummer. No improvement at all – but I need more testing to validate this.
@kimmonismus
-
Different GPT versions across devices and browsers
By
–
When I open ChatGPT in safari with my MacBook I have GPT-5.
When I open ChatGPT in brave on a Windows computer I have GPT-4o.
My own account is logged in.
In the iOS app again GPT-5. ?? -
Mixed Reactions to AI Model: Routing Issues and Inconsistent Performance
By
–
At least on X and Reddit, the mood seems to be rather disgruntled and disappointed. Biggest problem: routing. It just works incredibly poorly. In my testing as well. You should manually enable reasoning as needed. Additionally: it's better than GPT-4o and o3 – but often not
-

OpenAI’s Failed Presentation: Poor Charts and Quality Control Questions
By
–
With all due respect: But that really was the epitome of a fail. With the poorly made charts, OpenAI attracts ridicule. And I would love to know how that happened. I imagine that the presentation of the most important model was checked by someone, isn't it?
-
GPT-5 Impressive, But Google Genie 3 Delivers the Real Wow Factor
By
–
I really enjoy using GPT-5. But the real „wow“-effect so far was Genie 3. Google has given us a glimpse into the future – OpenAI an iterative update of their model.
-

Polymarket Reacts Negatively to GPT-5 Livestream Event
By
–
Looks like polymarket didn’t like the GPT-5 livestream
-

Most Important Benchmark Proving Exponential Development Progress
By
–
This is probably the most important benchmark and proof of ongoing exponential development,
-
Handwritten typos and German keyboard layout challenges
By
–
again some typos. That happens when you write by hand and dont use AI. And a German keyboard layout (ffs).
-

GPT-5 Release: Initial Thoughts and Livestream Analysis
By
–
GPT-5 & Livestream – my initial thoughts We have been waiting a long time for GPT-5. I still remember when the first rumors about GPT-5 began to circulate, back then under the code name Gobi, later Orion, but then it was released as GPT-4.5. Now it's finally here – and
-

GPT-5 Performance Trails Grok-4 Thinking in ARC-AGI-2 Benchmark
By
–
GPT-5 about 6% below Grok-4 thinking in ARC-AGI-2 GPT-5 ~10%
Grok-4 thinking~16%