Join us at our future events: luma.com/cerebras
@cerebras
-
Cerebras Event Draws Massive Crowd Seeking Speed
By
–
Packed room. Line around the block.
— Cerebras (@cerebras) 20 mars 2026
Why? Because everyone wants to go faster.
Appreciate everyone who showed up. pic.twitter.com/3ReBKlvRdHPacked room. Line around the block. Why? Because everyone wants to go faster. Appreciate everyone who showed up.
-

Cerebras Wafer Scale Advantage Over NVIDIA Groq Inference Chips
By
–
Problem solved. ✅ Andrew Feldman (@andrewdfeldman) NVIDIA's biggest GTC announcement was a $20 billion bet on the same problem we solved 6 years ago. Their next-gen inference chip – not available yet – has 140x less memory bandwidth than @cerebras. To run a single 2 trillion parameter model, you need 2,000+ Groq chips. On Cerebras, that's just over 20 wafers. Even paired with GPUs, Groq maxes out at ~1,000 tokens per second. We run at thousands of tokens per second today. And every day. In production now. Why? When you connect 2,000 chips together, every interconnect has latency. Every cable has overhead. It doesn't matter what your memory bandwidth is on paper if you're bottlenecked by the wiring between thousands of tiny chips. We solved this with wafer scale. One integrated system. Little interconnect tax. Jensen told the world that fast inference is where the value is. He’s right – it’s why the world’s leading AI companies and hyperscalers are choosing Cerebras. — https://nitter.net/andrewdfeldman/status/2034015373595672594#m
-
GPT-5.3-Codex-Spark: Three Real Workflows for Building
By
–
what can you build with gpt-5.3-codex-spark?@jxnlco from @OpenAI demos 3 real workflows — ones you can set up yourself inside the Codex app to help you spend less time on overhead and more time building.
— Cerebras (@cerebras) 18 mars 2026
00:09 – what is gpt-5.3-codex-spark?
00:25 – workflow 1: multi-agent… pic.twitter.com/ckpJREPJ3ewhat can you build with gpt-5.3-codex-spark? @jxnlco from @OpenAI demos 3 real workflows — ones you can set up yourself inside the Codex app to help you spend less time on overhead and more time building. 00:09 – what is gpt-5.3-codex-spark? 00:25 – workflow 1: multi-agent daily briefing from slack, drive & meets 01:06 – workflow 2: automated PR review 01:31 – workflow 3: real-time interactive coding 02:56 – what speed changes, and what's coming next
-
Link to article X from March 17 2026 by Cerebras
By
–
x.com/i/article/203369858369… [Translated from EN to English]
-
Cerebras Thanks Attendees at Morning Coffee Event
By
–
☕ Thanks to everyone who grabbed coffee with us this morning.
— Cerebras (@cerebras) 16 mars 2026
🟧 The big chip is just getting started.
See you around! pic.twitter.com/CTZdbI5GwW☕ Thanks to everyone who grabbed coffee with us this morning. 🟧 The big chip is just getting started. See you around!
-

AWS and Cerebras Partner for Order of Magnitude Faster Inference
By
–
We're teaming up with @cerebras to build the fastest possible inference. Coming soon to Amazon Bedrock, we’re delivering inference performance an order of magnitude faster than what’s available today by connecting AWS Trainium3 for compute-intensive prefill with Cerebras CS-3 to power decode. Learn more about the partnership. go.aws/3Pzcota
-
Cerebras Codex Spark Revolutionizes Coding with 1,200 Tokens Per Second
By
–
Our coding workflows were designed to accommodate slow inference. @OpenAI's Codex Spark powered by @cerebras changes the game.
— Cerebras (@cerebras) 12 mars 2026
Here's how we make the most out of 1,200 tokens per second, with @MilksandMatcha. pic.twitter.com/vv4a80wfFAOur coding workflows were designed to accommodate slow inference. @OpenAI's Codex Spark powered by @cerebras changes the game. Here's how we make the most out of 1,200 tokens per second, with @MilksandMatcha.
