heck, I am keeping my 8gb 3060 that I used to stream on for RAG purposes
MACHINE LEARNING
-
Compute usage: 1B ChatGPT vs 5M Codex users
By
–
Who is using more compute – 1b of ChatGPT users or 5m of Codex users?
-
Try Grok models on Cloudflare’s AI Gateway
By
–
Try Grok models on @Cloudflare's AI Gateway! https://t.co/YY511gTthP
— xAI (@xai) 3 juin 2026Try Grok models on @Cloudflare
's AI Gateway! -
Major GPT-Rosalind Upgrade for Enhanced Drug Discovery and Design
By
–
Major upgrade to GPT-Rosalind, with much better intelligence for drug discovery, analysis, design, and experimental workflows:
-
Harvey Legal Benchmark: 1250 tasks across 24 practice areas
By
–
We’ll see it more and more Btw Harvey Legal Benchmark is very broad more than « super focused ». 1,250 legal tasks, 24 legal practice areas, it was made to be one of the first large-scale realistic legal benchmark
-

This alone convinces enterprises to host LLMs on-premise
By
–
I mean, look at this, this alone is enough to get every enterprise out there into hosting their LLMs on-premise
-

Langchain guide on financial services agents from JPMorgan, Chime, Bridgewater
By
–
In our new guide, we break down lessons from @jpmorgan
, @Chime
, + Bridgewater on what it takes to bring agents into production in financial services, and how leading teams are building the operational foundation to ship with more confidence. Learn more: https://
info.langchain.com/guide/definiti
ve-guide-to-financial-services-agents-in-production
… -

LFM2/2.5 architecture: convolution is almost all you need
By
–
Oh fun! The LFM2/2.5 architecture is a joy to play with. Convolution is (almost) all you need.
-

World’s first heterogeneous disaggregated inference cloud shown live at ComputeX
By
–
The world's first heterogenous disaggregated inference cloud was just shown running live at ComputeX. VC2 — backed by a $3.5B compute commitment to SambaNova from @Vista_Equity & @cambiumcapital — brings three chips together in production for the first time:
– NVIDIA B200 GPUs