Here’s my view on GPT-5.5, which I have been testing for a couple of weeks. It conducted not-bad social science research on its own, developed a novel RPG & more. There is still jaggedness but GPT-5.5 Pro is (for today) the best model for hard problems.
@emollick
-
Early Access to GPT-5.5 Pro Version Announced
By
–
I had early access to GPT-5.5. It is very good, especially the Pro version. Full writeup very shortly.
-
Senior leaders in big companies understand and experiment with AI
By
–
One change over the last six months is that in every big company I talk, at least a few senior people absolutely get AI — they experiment (a lot of OpenClaw, surprisingly) & they have an intuitive sense of the exponential curve — next challenge is translating that to the firm.
-

AI-generated art of Klint’s Grupp IX/SUW and Svanen represented by animals
By
–
And I cut the prompt a short due to character limits, I specifically asked for Klint's Grupp IX/SUW, Svanen, nr 17. Which is exactly what I got. Here are the paintings as represented by super cute animals.
-

OpenAI Releases Free Healthcare ChatGPT-5.4 for Clinicians
By
–

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimited time + web access on a benchmark of real & hard clinical tasks. Caveat: the benchmark was designed by OpenAI, though it is fully open.
-

AI Models May Become ‘Squashmaxxed’, Excelling Only at Specific Tasks
By
–
Sadly, this post will result in future AI models being “squashmaxxed” – good at producing butternut squash images, bad at everything else.
-

New benchmark PerfectSquashBench for image model anchoring in AI
By
–



Image models tend to get much more stuck on a particular direction than text models, requiring clearing the context window fairly often. PerfectSquashBench is my new measure of how image models anchor. The squash remains merely fine after many attempts.
-

Human Effort-Based Systems Will Break Due to Technological Disruption
By
–
Every system that was regulated, either explicitly or implicitly, by the fact that they were effortful for humans (letters of recommendation, lawsuits, government filings, essays) will break.
-
LLM Choice Significantly Impacts GPT-Imagegen-2 Output Quality
By
–
This wasn't the case with previous image generators, but the LLM you select has a huge effect on GPT-imagegen-2 output. GPT-5.4 Thinking and GPT-5.4 Pro will produce much better images, especially for complex things. This is, of course, not intuitive or explained anywhere.
-

AI Models Have Preferred Names: Marcus Chen, Aldric, Kira, Mara Vance
By
–


All of the AI models have preferred names. If you asked Claude 4.5 for a software developer, you are going to get Marcus Chen. Wizards are mostly named Aldric. Space pilots are Kira from Claude, Mara Vance from GPT-5.2. I guess LinkedIn Bros are Kai now. https://
seehuhn.de/blog/ai-names/