This news comes hours after @NewYorker published its investigation detailing the various ways AI experts warn OpenAI hasn't been taking AI safety seriously enough.
GENERATIVE AI
-
Benchmarks and hallucinations in reasoning models clarified
By
–
that’s one particular benchmark that came up; not a universal claim by me (or anyone else AFAIK). compare “Gary mentioned a particular benchmark” that didn’t include reasoning models (true) with “no studies have indicated hallucinations within reasoning models” (false). you
-
AI Agents Transform the Web into Autonomous Platform
By
–
AI is changing the Internet into a platform where you can build literally anything. Travel site? Check. Weather site? Check. An app to watch your favorite sports team? Check. Automatic shopping? Check. And millions of other things. Arlan lays out how to turn the Web into a file system. A new web is here. The autonomous web. Agents are like autonomous vehicles. Get in a Tesla and it will take you to In-N-Out. Or anywhere. It turns the world into a thrill park ride. Same for your agents on the Web. They are like little autonomous vehicles that can do anything. Why do you think the OpenClaw got the nerds so excited? They are building systems to do millions of things automatically on the Web. Where's @timoreilly? He wrote the Web 2.0 memo that gave that its name. Is this Web 4.0? I'd rather call it the autonomous web. How about you? Arlan (@arlanr) x.com/i/article/203981724469… — https://nitter.net/arlanr/status/2041215978957389908#m
→ View original post on X — @scobleizer, 2026-04-06 20:29 UTC
-
Elon’s AI Expected to Generate Scripts and TV Shows by Year-End
By
–
Yeah, and soon Elon says it will be able to read lists, create scripts, and create TV shows and other stuff from the lists. I expect that by the end of the year.
-
Codex OSS Update: Which Projects Should We Support Next
By
–
Codex for OSS Update – Which projects should we support next? 🙂 DMs open!
-
Stabilizing Video from Running Animals: New Petpin Pipeline Breakthrough
By
–
Stabilizing video from a camera on a running animal turns out to be brutally hard.
— Ark Baltser (@arkslife) 6 avril 2026
Traditional stabilization breaks pretty quickly.
We're starting to crack it.
Before → After from our latest Petpin pipeline. pic.twitter.com/L4NWftjfxCStabilizing video from a camera on a running animal turns out to be brutally hard. Traditional stabilization breaks pretty quickly. We're starting to crack it. Before → After from our latest Petpin pipeline.
→ View original post on X — @scobleizer, 2026-04-06 20:17 UTC
-
LLMs Auto-Regressive Nature: Beyond AI Hype
By
–
Guessing you're referring to the story and not my tweet, but totally agree that everyone should remember LLMs are by design auto-regressive. (Shocking so many still don't.) AI companies should stop over-hyping LLMs and start explaining how they actually work and where they're
-

Anthropic Head of Growth: AI Requires More PMs, Not Fewer
By
–
Narrative violation: Anthropic's Head of Growth says we'll need more PMs, not fewer.
— Lenny Rachitsky (@lennysan) 6 avril 2026
"While PMs and designers are getting leverage from AI, engineering is getting the most leverage right now. If you think about a default team with 5 engineers, 1 designer, 1 PM—with Claude Code,… https://t.co/dnF0CeuudT pic.twitter.com/LisCpN30SHNarrative violation: Anthropic's Head of Growth says we'll need more PMs, not fewer. "While PMs and designers are getting leverage from AI, engineering is getting the most leverage right now. If you think about a default team with 5 engineers, 1 designer, 1 PM—with Claude Code,
-
Base LLMs Fail at Math, LRMs Make Progress
By
–
Paper below tested a variety of base LLMs (no TTA) on generalization-focus math problems and found that they can't reason and can't do math. All true… but the fact that base LLMs have zero fluid intelligence, while extremely controversial back in 2024, is now well established. An interesting experiment here would have been to try current LRMs on the same problems and measure the delta. I bet latest LRMs can solve most of these problems. arxiv.org/abs/2604.01988 [Translated from EN to English]
-
Can AI Build a Simple Single-Purpose App Correctly?
By
–
I’m talking about a very simple app that does just one thing and does it correctly. Something that even *I* could develop if someone told me what it is. If AI can’t do it then IMHO fails a very very straightforward test of intelligence.
