I am not kidding, now is the time more than ever to hunt an RTX 3090 and learn how to run Qwen 3.5 27B
AI HARDWARE
-

@garymarcus — 2026-06-26
By
–
“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as railways survived the 19th-century railway bust. However, this fails to reckon with the reality of depreciation (few pieces of silicon hold
-
Nvidia’s biggest moat is not CUDA or GPUs
By
–
It's actually way dumber than that. CUDA is not even the biggest Nvidia moat. And neither are the GPUs. Both of those could be completely commoditized tomorrow and it would hardly have any impact on Nvida's AI infra dominance.
-
Almost added home GPU rig but local LLM lacks cloud quality
By
–
So I almost added
– Home GPU rig And it might be nice for self-sufficiency but again I still don't think local LLM stuff even comes close to cloud either in quality or performance (speed) or cost Making your home self-sufficient is nice though, I have:
– 2x Tesla Powerwall (27 -
GLM 5.2 MoE, NVFP4 467GB, DGX Station memory, offloading works
By
–
GLM 5.2 is an MoE, NVFP4 is 467 GB, and the DGX Station comes with 496GB LPDDR5X + 252GB HBM3e GPU memory With the right offloading formula, it should work
-
NVIDIA-accelerated AI aids PYLER in brand safety for advertisers
By
–
Every day, millions of videos compete for advertising dollars. Ensuring brands appear alongside the right content requires AI that can understand context at scale.
— NVIDIA (@nvidia) 25 juin 2026
PYLER is helping advertisers improve brand safety and campaign performance with NVIDIA-accelerated AI that analyzes… pic.twitter.com/9xSDjj9e9gEvery day, millions of videos compete for advertising dollars. Ensuring brands appear alongside the right content requires AI that can understand context at scale. PYLER is helping advertisers improve brand safety and campaign performance with NVIDIA-accelerated AI that analyzes
-
User finds local LLMs painfully slow on single RTX 5090
By
–
And yes I've tried local LLMs but with just 1x RTX 5090 it's painfully slow and useless
-
Gemma 4 prioritizes local on-device intelligence across hardware classes
By
–
Gemma 4 is best in class at each hardware class, not designed to compete on server side frontier intelligence like GLM, it’s designed to enable local on device intelligence without needing advanced hardware
-
SambaNova demonstrates first disaggregated inference cloud for AI agents
By
–
Disaggregated Inference Is Live! At COMPUTEX, we demonstrated the world's first disaggregated inference cloud for AI agents.
GPUs for prefill. RDUs for decode. CPUs for orchestration. The result: faster agent workflows and better economics than homogeneous infrastructure. -
Revolut’s transaction foundation model accelerated by NVIDIA and Nebius
By
–
Revolut built a transaction foundation model – accelerated by the NVIDIA full-stack platform on Nebius – to improve fraud detection, product recommendations, and other use cases in financial services.
— NVIDIA (@nvidia) 25 juin 2026
The results:
📈 2.3x better credit risk accuracy
⚡ Up to 5x higher training… pic.twitter.com/vAtbVJXfPZRevolut built a transaction foundation model – accelerated by the NVIDIA full-stack platform on Nebius – to improve fraud detection, product recommendations, and other use cases in financial services. The results: 2.3x better credit risk accuracy Up to 5x higher training