Wall Street is watching frontier model benchmark porn while enterprises quietly die waiting for capabilities that actually ship. MMLU and HLE were very useful for about a nano-second. But now they are just boring. We don’t need to know how well an LLM can do on a test. We need
TECHNOLOGY
-
New AI coming tonight, author to watch
By
–
New AI's coming tonight. I'll be watching. https://t.co/KH4zHDWrDt
— Robert Scoble (@Scobleizer) 1 juin 2026New AI's coming tonight. I'll be watching.
-
Deep dive into Gaussian Splatting: exploring advanced techniques and applications.
By
–
go deeper into gaussian splatting:
-
Bilawal Sidhu praises Lichtfeldstudio’s advanced NeRF methods
By
–
def one of the big leaps since the cloudy NeRF days — @lichtfeldstudio really does have the best methods an dials rn!
-

Teaching Codex to automate QA testing of OpenClaw via VNC and browser
By
–
Been teaching codex to be my QA assistant. For every commit it creates a user-test scenario and uses webVNC (crabbox), computer/browser use (peekaboo/mcporter) to test OpenClaw like a user/QA person would. This runs in the background and opens PRs with fixes.
-
Generational upgrade of evals/analytics into continual learning platforms
By
–
every evals/analytics startup is going through a onetime generational upgrade into a continual learning platform in 2026 many will fail but as always the tasteful ones win
-
Build Autonomous Claude Code Routines in 15 Mins with Anthropic Engineer’s Guide
By
–
This Anthropic engineer reveals how to build rad Claude Code Routines in 15 mins.
— Charly Wargnier (@DataChaz) 31 mai 2026
The result?
A 24/7 autonomous coding partner.
Along with @zodchiii's guide, you'll automate 50% of your workload with Claude Code.
No fluff. Just the playbook from the ppl who actually built it. https://t.co/b9b0dzFvb2 pic.twitter.com/LG0Vro1Gt5This Anthropic engineer reveals how to build rad Claude Code Routines in 15 mins. The result? A 24/7 autonomous coding partner. Along with @zodchiii
's guide, you'll automate 50% of your workload with Claude Code. No fluff. Just the playbook from the ppl who actually built it. -
Rational conversation on AI future with Benedict Evans
By
–
A rational conversation on where AI is actually going with @benedictevans For 20+ years, Benedict has been one of the clearest, most reliable thinkers on where technology is heading, and how it'll impact our lives. He was @a16z
's resident "thinker" for 5+ years, and has spent -
Solving a multi-agent management problem with Google DeepMind
By
–
It's been almost 1 and a half years that I can't solve a multi-agent cohort management problem and now I just came across a library from @GoogleDeepMind released 2 months ago that offers an ultra elegant approach. Damn, this company is really the boss.
-
Human imperfection compared to machine capabilities in various tasks
By
–
If we are so perfect, why are we so bad at arithmetics, computing integrals symbolically, playing chess, go, or poker?
Machines are already more "perfect" than us at these and many other tasks. (Also, why can we die from diseases if we are so perfect?)
