Depends on the use-case… for many simple consumer use-cases, sure, but for anything that requires thinking, iteration, etc. t/s is insanely important. Extremely fast inference from @GroqInc or similar cuts down a 10 minute thinking workflow into just a few seconds. Or allows
COMPUTING
-
Formalizing Hofstadter’s Strange Loop Beyond Recursion
By
–
I agree: Hofstadter’s concept of the strange loop is not a specification that can be directly implemented, and recursion can be vexing but is not necessarily strange. Hofstadter is pointing at something interesting that is yet to be formalized.
-
AI Factories Boost Efficiency Through Accelerated Computing Technology
By
–
AI factories can use accelerated computing to increase tokens per watt, optimizing AI performance and energy efficiency.
-
Optimizing AI Output Across Different Operating Configurations
By
–
It visualizes the best output for different operating configurations.
-

AI factories redefining modern infrastructure economics
By
–
AI factories are redefining the economics of modern infrastructure. AI factories help enhance three key aspects of the AI journey: Data ingestion Model training High-volume inference
-

Siemens Combines AI and Digital Twins for Manufacturing Innovation
By
–
By combining AI, digital twins, and accelerated computing, @Siemens is helping manufacturers: Speed up product development Boost factory productivity Collaborate in real time Enhance design reviews Learn more: https://
nvda.ws/4luFDIU -
AI Lab Offers Unlimited GPUs Per Researcher Hiring
By
–
My AI lab has infinite GPUs per researcher. Join now, before we hire anyone.
-

Advanced AI Platform for Professionals: Deep Reasoning and Extended Context
By
–
Designed for pros Developers
Researchers
Data Academics
Teams needing deep reasoning
Large-context processing (128k in-app / 256k tokens API)
Early access to cutting-edge video / coding tools. -
Open-Source Frontier Model Reaches 185 Tokens Per Second
By
–
We officially have a near-frontier open-source model running on @GroqInc at 185 tok/s.
— Matt Shumer (@mattshumer_) 15 juillet 2025
It’s only going to get faster from here.
This is going to open up a lot of opportunities. https://t.co/P8Cs1SQgCg pic.twitter.com/PwBhDjQdEzWe officially have a near-frontier open-source model running on @GroqInc at 185 tok/s. It’s only going to get faster from here. This is going to open up a lot of opportunities.
-

K2 Model: Efficient Token Usage Without Reasoning
By
–
Remember: K2 is *not* a reasoning model. And very few active tokens in the MoE. So it's using less tokens, *and* each token is cheaper and faster.