Just deployed additional speed optimizations for @cognition SWE-1.5. The fastest measured request was an eye watering 1,881 token/s per Grafana dashboard.
@cerebras
-
Frontier Models vs GPT Codex: Speed to Success Comparison
By
–
Even frontier models fail all the time. The difference is we fail in 15 sec and with one more prompt you get the right answer. Time to success = 30 sec. On GPT Codex it takes 22min just to find out it failed.
-
Cerebras Introduces Code Platform for AI Development
By
–
https://
cerebras.ai/blog/introduci
ng-cerebras-code
… -

SWE-1.5 fastest coding agent released by Cognition Cerebras
By
–
Today, @cognition released SWE-1.5 – the world’s fastest coding agent, powered by Cerebras. SWE-1.5 achieves frontier-level coding ability, comparable to Sonnet 4.5 and surpassing GPT-5. Cerebras and Cognition engineers worked hand in hand over the past few weeks, training a
-

OpenAI Releases GPT-OSS-Safeguard Open-Weight Model
By
–
GPT-OSS-Safeguard from @OpenAI is here. Open-weight, safety-tuned, transparent reasoning. Now available in private preview at Cerebras speeds https://
cerebras.ai/build-with-us -

SWE-grep: ML-Powered Code Search and Explanation Tool
By
–
SWE-grep is truly some inspired ML from @cognition
. Take a 1M line codebase like React. With multiple fast inference calls on Cerebras, it can fetch & explain relevant code in seconds. Here are a few measurements taken on React, Vercel, PyTorch repos. -
Cerebras Powers Cognition’s Code Retrieval in Windsurf
By
–
Cerebras is now powering Cognition's latest code retrieval models directly in @windsurf
— Cerebras (@cerebras) 16 octobre 2025
Context retrieval has been one of the biggest bottlenecks in agentic coding. When you ask an agent to work on a large codebase, it can spend 60% of its time just searching for relevant files.… https://t.co/fAMRponTtoCerebras is now powering Cognition's latest code retrieval models directly in @windsurf Context retrieval has been one of the biggest bottlenecks in agentic coding. When you ask an agent to work on a large codebase, it can spend 60% of its time just searching for relevant files.
-
The Greatest AI Chip Ever Made
By
–
the greatest AI chip ever made https://
arxiv.org/html/2503.1169
8v1
…

