Sonnet 4.5's character has also improved significantly. Its responses are more to the point with less LLM fluff. We've tuned the model across a bunch of dimensions – sycophancy, deception, prompt injection robustness. You'll notice it in how the model actually talks to you.
GENERATIVE AI
-

Sonnet 4.5 Achieves 82% on SWE-bench with Test-Time Compute
By
–
Sonnet 4.5 hits 82% on SWE-bench with test-time compute. We're adding another dot to what's been a pretty clean exponential toward saturating SWE-bench as a benchmark. At this rate we'll need new evals soon, but it's still a useful signal that the capability curve is holding.
-

Claude Sonnet 4.5: Advanced Coding Model with Superior Design
By
–
Introducing Claude Sonnet 4.5. The best coding model in the world combined with the best character of any model I've seen.
-

Anthropic to release Claude Sonnet 4.5 soon
By
–

BREAKING : Anthropic is about to release Claude Sonnet 4.5 soon! SOTA on SWE bench
-

Accenture exits 11,000 staff unable to retrain for AI age
By
–
As growth softens, expect more “AI is taking our jobs” headlines. Accenture says it will exit about 11,000 people – described as staff who “cannot be retrained for the age of AI.” The cause is uncertain, but the scale isn’t: since 2020, Accenture has added ~273,000 roles (+54%).
-

DeepSeek’s Sparse Attention Enables Massive Price Reduction Economics
By
–
>50% price cut across the board *MAKES DeepSeek money*
— swyx 🐣 (@swyx) 29 septembre 2025
put another way, DeepSeek's Sparse Attention allows DeepSeek to serve 1.8m OUTPUT tokens for the same price as it used to serve 128k output.*
here's my extremely scientific calculation (i also eyeballed it and did it on a… https://t.co/rDNwafshwQ pic.twitter.com/5FpN3V2gQM>50% price cut across the board *MAKES DeepSeek money* put another way, DeepSeek's Sparse Attention allows DeepSeek to serve 1.8m OUTPUT tokens for the same price as it used to serve 128k output.* here's my extremely scientific calculation (i also eyeballed it and did it on a
-
AI Impact on Programming Careers Reassurance Note
By
–
More on my blog, including this hopefully reassuring note for anyone afraid of the impact this will have on their career as a programmer https://
simonwillison.net/2025/Sep/29/ar
min-ronacher-90/
… -
AI-Generated Code in Infrastructure: Armin’s Credible 90% Achievement
By
–
Most of the "90% of code written by AI" claims come from vendors selling AI tools and lack credibility as a result Armin (creator of Flask, Jinja, Click) is different – when he says 90% of a new infrastructure project he's building was AI generated it's worth paying attention
