It scored a 62/100 on my Senior Engineer benchmark, a new record! Worked for almost a day:
AGI
-
AI Model Release Furor: Don’t Need to Switch Providers or Declare a Winner
By
–
Among all the model release furor, important to note that people don’t need to switch providers or declare a winner every time a new model is released. Especially as Opus 4.7 is a good model too! (Especially since adaptive thinking got better)
-
AGI-Pilled Product Management Skills for 2026
By
–
Claude Code's Head of Product: "The hardest PM skill right now is how to be the right amount of AGI-pilled." https://t.co/IebAy2xB2k pic.twitter.com/03Y9AmmRu7
— Lenny Rachitsky (@lennysan) 23 avril 2026Claude Code's Head of Product: "The hardest PM skill right now is how to be the right amount of AGI-pilled."
-
GPT-5.5 Pro Testing: Social Science Research, RPG Development, and Hard Problems
By
–
Here’s my view on GPT-5.5, which I have been testing for a couple of weeks. It conducted not-bad social science research on its own, developed a novel RPG & more. There is still jaggedness but GPT-5.5 Pro is (for today) the best model for hard problems.
-

Unpredictable AGI may resist control, diverse AI safer
By
–
Unpredictable AGI may resist full control, making diverse #AI safer
by Gaby Clark @TechXplore_com Learn more: https://
bit.ly/48empCF #ArtificialIntelligence #MachineLearning #ML -

GPT-5.5 performance on ARC-AGI-2 benchmark
By
–
GPT-5.5 just hit 85% on ARC-AGI-2. Somewhere, Yann LeCun is explaining why this still doesn't count. LLMs keep climbing a tall tree toward the moon
-
GPT-5.5 Live Launch Announced by OpenAI
By
–
GPT-5.5 Live with Every https://
x.com/i/broadcasts/1
AxRnaLmZYPxl
… -

User claims AI model ‘Claude’ was intentionally degraded
By
–

I can't believe we were right Claude was dumbified on March 4, just when we noticed!
-
Early Access to GPT-5.5 Pro Version Announced
By
–
I had early access to GPT-5.5. It is very good, especially the Pro version. Full writeup very shortly.
-

Smaller Token Count Improves Agentic Programming Model Performance
By
–
Once again, in agentic programming itself the same pattern. The model achieves better performance using a smaller number of tokens.