I was full on promoting Opus as best model for “Clawdbot”. Luckily that changed and GPT 5.5 is now the best model based on our internal benchmarks.
AI
-
Reconstructing Software Engineering for AI-Driven Coding
By
–
Reconstructing software engineering around AI is going to take work (even as the ability of AI to code increases at a rapid rate). Organizations are ideally spending tokens for two things:
1) building stuff
2) experiments to figure out best practices (which involves failure) -
YOLOE: Open-Vocabulary Detection at YOLO Speed
By
–
Object detection is shifting from "models that recognize fixed categories" to "models that understand concepts described in language."
— Satya Mallick (@LearnOpenCV) 29 mai 2026
YOLOE delivers open-vocabulary detection at full YOLO speed — text module fused into the head, zero runtime overhead.
Full tutorial + code:… pic.twitter.com/0yRPYv5jRUObject detection is shifting from "models that recognize fixed categories" to "models that understand concepts described in language."
YOLOE delivers open-vocabulary detection at full YOLO speed — text module fused into the head, zero runtime overhead.
Full tutorial + code: -
Harness Profiles for Multi-Model Prompt Optimization
By
–
Different models need different prompts, sometimes tools “Harness profiles” are how we do that in deepagents
-
2026 budget spent early not notable due to unknown Claude Code in 2025
By
–
I don't think "we spent our 2026 budget in 4 months" is particularly notable given that the budget would have been set in 2025 when nobody knew how good Claude Code et al were about to get as-of January
-

Updated post on product market fit with AI failure stories note
By
–
Just updated my post here to include the "harder to justify" note https://
simonwillison.net/2026/May/27/pr
oduct-market-fit/#the-ai-failure-stories-around-this-are-pretty-thin
… -

OpenAI Codex introduces Side Conversions in ChatGPT for side questions
By
–


OPENAI : Codex in ChatGPT now supports Side Conversions, allowing users to ask side questions without disrupting the main thread. /Side testing
-
System built with Codex due to Claude’s mistakes
By
–
nah, it’s built with codex. Claude makes too many mistakes.
-

Model Release Cycles: Anthropic & OpenAI Speed Advantage
By
–
The model release cycles from Anthropic & OpenAI are genuinely insane, previously we had 6-12 months between updates, now models are released every 1.5 months. This is an under-appreciated reason why Anthropic & OpenAI are in the lead. Google's releases are not as fast,



