
Real world example: raising token limits from 3M to 10M tripled the amount of work that Codex could do independently on cybersecurity tasks. From 3.1 hours to 10.5 hours.

By
–

Real world example: raising token limits from 3M to 10M tripled the amount of work that Codex could do independently on cybersecurity tasks. From 3.1 hours to 10.5 hours.

By
–
Gemma 4 is #1 on @huggingface!
→ View original post on X — @demishassabis, 2026-04-05 21:55 UTC

By
–

Unappreciated fact is the second scaling law does not seem to completely plateau in many tasks: throw more tokens at a reasoning AI model and get better answers, especially with a simple harness. Benchmark performance is actually limited by token usage. https://
open.substack.com/pub/joelbkr/p/
many-benchmarks-scores-would-appear?r=i5f7&utm_medium=ios
…
By
–
Iran Threatens "Complete And Utter Annihilation" Of OpenAI's $30BN Stargate Data Center In Abu Dhabi zerohedge.com/geopolitical/i… [Translated from EN to English]
→ View original post on X — @olivierrimmel, 2026-04-05 21:50 UTC
By
–
I wouldn't have a policy that says you're not allowed to use "claude -p"

By
–
PhAIL just launched — the first robotics benchmark measuring units per hour and mean time between failures, not "success rate." First time I've seen the research community measure robots the way a factory operator would. The gap between those two measurements is where most deployments quietly fail. [Translated from EN to English]
→ View original post on X — @ken_goldberg, 2026-04-05 21:42 UTC
By
–
Connected devices are reshaping plant operations. Anyone running into this yet? @IIoT_World @CRudinschi @agentic_factory @fabot70 @gp_pulipaka @jblefevre60
By
–
If you're not using Keras with JAX you're ngmi
By
–
Amazing what AI can do when unions don’t get in the way. reuters.com/technology/ai-is…
→ View original post on X — @pmddomingos, 2026-04-05 21:23 UTC