“Test-time compute” is such a dumb name for “compute”. So many dumb names for things in AI.
AI
-
Open source weights and 1m context win
By
–
Yeah, but opensource and weights and 1m context. So I’d say it’s a win
-
Grok, Cosmos, and World Models: The Video Agent Moment
By
–
Grok Imagine’s Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://
latent.space/p/video-agents @EthanHe_42
, former @xai world model lead and @nvidia Cosmos researcher, explains why AI video may follow the same path as coding agents, how Grok -
DellTechWorld 2026: H2O.ai CEO and Dell CTO discuss AI scaling, governance, cost
By
–
At #DellTechWorld 2026, http://
H2O.ai Founder & CEO @srisatish joined @DellTech
' CTO Satish Iyer and @theCUBEto discuss what it takes to operationalize AI at scale — from infrastructure and governance to cost control, sovereignty, and real-world business outcomes. -

MiniMax M3: 59% SWE-Bench Pro, ahead of GPT-5.5 and Gemini 3.1 Pro
By
–

MiniMax just dropped M3! It hits 59% on SWE-Bench Pro, edging out GPT-5.5 (58.6%) and beating Gemini 3.1 Pro (54.2%). Trails Opus 4.7 on coding, but leads it on autonomous browsing at 83.5% on BrowseComp. First open model to pack frontier coding, a 1M-token context, and native
-
M3: MiniMax’s latest model with 1M-token context and agentic reasoning
By
–
M3 is MiniMax's latest model: 1M-token context, native multimodality, and agentic reasoning, built on their Sparse Attention architecture. Try it in Atomic Chat
-
MiniMax M3 live in Atomic Chat — HTML game
By
–
MiniMax M3 is now live inside Atomic Chat 👀
— 🚨 AI News | TestingCatalog (@testingcatalog) 1 juin 2026
Atomic tested M3 on a task to read a hand-drawn napkin sketch, write the game logic, build the UI, and ship a playable HTML platformer in one pass.
All this for $0.028 🤖 https://t.co/KfXqhytXNa pic.twitter.com/RyKjPioVH8MiniMax M3 is now live inside Atomic Chat. Atomic tested M3 on a task to read a hand-drawn napkin sketch, write the game logic, build the UI, and ship a playable HTML platformer in one pass. All this for $0.028.
-
Demucs produced at FAIR-Paris by @honualx
By
–
Demucs was produced at FAIR-Paris by @honualx and collaborators.
-
Agents need memory before tools, not expensive goldfish
By
–
We gave agents tools before we gave them memory.
— God of Prompt (@godofprompt) 1 juin 2026
That was backwards.
Now they can browse, code, email, book meetings, hit APIs…
…but still wake up tomorrow with no idea what they learned yesterday.
That is not an agent stack. That is a very expensive goldfish with… https://t.co/k1DqTtViOb pic.twitter.com/vH1kBgbqNtWe gave agents tools before we gave them memory. That was backwards. Now they can browse, code, email, book meetings, hit APIs… …but still wake up tomorrow with no idea what they learned yesterday. That is not an agent stack. That is a very expensive goldfish with
-

Codex working nonstop since 11am, feels like an AI employee
By
–
I launched Codex on a goal at 11am, it's 5pm and it's still working. I've never felt so strongly the feeling of having an AI working for me. Like, it's doing a normal workday right now. I'm working on something else while it…