
BREAKING : Windsurf has released a new SWE-1.5 model that delivers "near-SOTA coding performance" at significantly higher speeds. Falcon has been revealed

By
–

BREAKING : Windsurf has released a new SWE-1.5 model that delivers "near-SOTA coding performance" at significantly higher speeds. Falcon has been revealed
By
–
it greatly REDUCED them overall when I used ChatGPT thinking.
By
–
why does microsoft copilot look like a potato can you answer this
By
–
Available in public beta today on the Claude API and on Google Cloud’s Vertex AI, with Amazon Bedrock coming soon. Docs here:
By
–
We just added thinking block preservation in the Claude API. You can now control how thinking blocks are managed in your context window, resulting in more cache hits and lower costs.

By
–
In general, Claude Opus 4 and 4.1, the most capable models we tested, performed best in our tests of introspection (this research was done before Sonnet 4.5). Results are shown below for the initial “injected thought” experiment.

By
–
We also found evidence for cognitive control, where models deliberately "think about" something. For instance, when we instruct a model to think about "aquariums” in an unrelated context, we measure higher aquarium-related neural activity than if we instruct it not to.
By
–
However, it doesn’t always work. In fact, most of the time, models fail to exhibit awareness of injected concepts, even when they are clearly influenced by the injection.

By
–
GPT-OSS-Safeguard from @OpenAI is here. Open-weight, safety-tuned, transparent reasoning. Now available in private preview at Cerebras speeds https://
cerebras.ai/build-with-us
By
–
BREAKING 🚨: Cursor got a major upgrade to version 2.0 with a multi-agent system and its own coding model!
— 🚨 AI News | TestingCatalog (@testingcatalog) 29 octobre 2025
– Cheetah is now Composer-1, and it is 4 times faster than other frontier models.
– The same prompt can run across multiple models
– Built-in browser for testing
– Voice… https://t.co/QfuyuSjpNZ pic.twitter.com/MbNFILsSOV
BREAKING: Cursor got a major upgrade to version 2.0 with a multi-agent system and its own coding model! – Cheetah is now Composer-1, and it is 4 times faster than other frontier models.
– The same prompt can run across multiple models
– Built-in browser for testing
– Voice