building things at 1,000 tokens/s
LLMS
-

Token Cost Depends on AI Infrastructure Expenses
By
–
The cost of each token depends on the cost and the overall token output of AI infrastructure.
-

AI Scaling: Solving Token Economics at Scale
By
–
As AI scales, the big question often is: How can you afford more tokens? The answer: better tokenomics.
-

The Raven Paradox: AI and Machine Learning Logic Explored
By
–
The Raven Paradox – Probably Overthinking It https://
buff.ly/sacgzyL
#AI #MachineLearning #DeepLearning #LLMs #DataScience -
Developers’ Request for Faster Codex Before Codex-Spark
By
–
Me talking to developers two weeks ago who wanted a faster Codex.
— Romain Huet (@romainhuet) 12 février 2026
That was before Codex-Spark! ✨ pic.twitter.com/1NLdbWo1IaMe talking to developers two weeks ago who wanted a faster Codex. That was before Codex-Spark!
-
Cerebras OpenAI Codex Spark Advanced AI Hardware Innovation
By
–
https://
cerebras.ai/blog/openai-co
dexspark
… -
OpenAI Codex-Spark: 1000 tokens/s code generation speed
By
–
OpenAI Codex-Spark powered by Cerebras
— Cerebras (@cerebras) 12 février 2026
You can now just build things faster—at 1,000 tokens/s. pic.twitter.com/4qxxVwZv4dOpenAI Codex-Spark powered by Cerebras You can now just build things faster—at 1,000 tokens/s.
-
Discussion on LLM inference tokens per second
By
–
can you feel more than 1000s of tokens per second, anon!? ✨
— Vaibhav (VB) Srivastav (@reach_vb) 12 février 2026
codex -m gpt-5.3-codex-spark pic.twitter.com/6UJzWE6qqjcan you feel more than 1000s of tokens per second, anon!? codex -m gpt-5.3-codex-spark
-
OpenAI launches GPT-5.3-Codex-Spark on Cerebras infrastructure
By
–
BREAKING 🚨: OpenAI released GPT-5.3-Codex-Spark, a new, faster model powered by @cerebras infrastructure!
— 🚨 AI News | TestingCatalog (@testingcatalog) 12 février 2026
Available to Pro users only as a research preview. https://t.co/DVrKJcLt63 pic.twitter.com/6JoarpB8zrBREAKING : OpenAI released GPT-5.3-Codex-Spark, a new, faster model powered by @cerebras infrastructure! Available to Pro users only as a research preview.
-
GPT-5.3-Codex-Spark: Real-Time Coding Model with 1000+ Tokens Per Second
By
–
Hello GPT-5.3-Codex-Spark! ✨
— Romain Huet (@romainhuet) 12 février 2026
Our first real-time coding model. It is… FAST. 1,000+ tokens per second.
Once you experience latency this low, it’s hard to go back. This is an exciting first milestone in our partnership with @Cerebras. pic.twitter.com/ieuDkJ1qLzHello GPT-5.3-Codex-Spark! Our first real-time coding model. It is… FAST. 1,000+ tokens per second. Once you experience latency this low, it’s hard to go back. This is an exciting first milestone in our partnership with @Cerebras
.