Announced at #GoogleCloudNext Enterprises can now unlock the full potential of #AgenticAI with the latest Gemini models now available on Google Distributed Cloud with NVIDIA Confidential Computing on NVIDIA Blackwell infrastructure. Learn more
AI HARDWARE
-

DataRobot NVIDIA Enable Quick Enterprise AI App Deployment
By
–
With DataRobot and @NVIDIA AI Enterprise, AI teams can quickly develop and deploy AI apps without the hassle of assembling solutions from scratch. Learn more: https://
bit.ly/4jjyk5q #agenticAI #enterpriseAI #AIapps #AIagents -
US Pauses Tariffs 90 Days Except China at 125%
By
–
Slow the madness, tariffs are officially paused for 90 days. except for China that now has 125% tariffs.
-
UC Researchers Achieve Real-Time Brain-to-Speech Synthesis Breakthrough
By
–
Work led by BAIR students @KayloLittlejohn and @CheolJunCho advised by BAIR faculty @GopalaSpeech "…made it possible to synthesize brain signals into speech in close to real-time." https://
dailycal.org/news/campus/re
search-and-ideas/uc-researchers-make-breakthrough-on-brain-to-speech-device/article_a438d4e4-1973-45fd-839c-6cd865906586.html
… via @dailycal -

Llama 4 Maverick fastest inference 655 tokens per second
By
–
I feel the need, the need for speed! Llama 4 Maverick from @AIatMeta is now available on SambaNova Cloud & it's the fastest inference verified by @ArtificialAnlys at 655 t/s. $0.50 / million input tokens & $2.00 per million output tokens. On SambaNova Cloud
-

Nvidia and Semiconductor Firms Face Tariff Uncertainty Ahead
By
–
Wall Street’s Best Ideas: Get ready for even more confusion The coming earnings season is going to have a lot of half-hearted, confused forecasts from management at semiconductor companies, including Nvidia, as they struggle to gauge tariffs’ ultimate impact. One analyst is
-

Metis M.2 Powers Offline AI Chatbot on Arduino Portenta X8
By
–
Metis M.2 Meets Arduino Portenta X8 Check out the power of @AxeleraAI
’s Metis PCIe on the @arduino Portenta X8 at #ArduinoDay! Fully offline AI chatbot demo—no cloud needed. Now in Early Access: https://
eu1.hubs.ly/H0j8JM40 #EdgeAI #AIinference #Metis #Arduino #OfflineAI -

Understanding GPU Architecture Fundamentals for AI Systems
By
–
Fundamentals of GPU Architecture https://
buff.ly/Zc6p5ks
#AI #MachineLearning #DeepLearning #LLMs #DataScience -
Google Introduces Ironwood TPU for Inference Era
By
–
Introducing Ironwood, the first TPU built for the age of inference, and the timing could not be better : ) – Ironwood perf/watt is 2x relative to Trillium, 6th gen TPU
– Ironwood offers 192 GB per chip, 6x that of Trillium
– 4.5x faster data access -
Exponential AI Demand Requires Investment Across Full Stack
By
–
The demand for AI is on an exponential, we need to continue investing at every layer of the stack if we are going to be able to service the demand for AI compute efficiently and at humanity scale. TPU's are so wonderful : )