This is a sad day for the AI community… I'm working on a benchmark we will soon release… and these are early results gathering pretty much all models out there, closed and opened. Red = closed source models
y axis = Elo score
x axis = release date
size = task cost … and
TECHNOLOGY
-

Early benchmark results show closed source AI models leading
By
–
-
AI in Stores: Instacart’s Caper Carts with NVIDIA Jetson
By
–
Real grocery stores are hard on AI: changing shelves, spotty Wi-Fi and shoppers on the move.
— NVIDIA (@nvidia) 24 juin 2026
On the NVIDIA AI Podcast, @instacart 's David McIntosh shares how Caper Carts use NVIDIA Jetson and edge AI to recognize items in store. pic.twitter.com/oPXNhjJ5uUReal grocery stores put AI to the test: constantly changing shelves, unstable Wi-Fi, and moving customers. In the NVIDIA AI Podcast, Instacart’s David McIntosh shares how Caper Carts use NVIDIA Jetson and AI.
-

Qwen-AgentWorld: Linguistic World Models for General Agents
By
–
Qwen-AgentWorld Linguistic World Models for General Agents
-
Sakana AI CEO David Ha interviewed on TBS CROSS DIG about vision and products
By
–
Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical vision, product launches including Sakana Fugu, Japan’s AI strategy, and his hopes for Japanese society. Watch the interview. https://t.co/jfRtraGpw4
— Sakana AI (@SakanaAILabs) 24 juin 2026Sakana AI CEO David Ha (
@hardmaru
) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical vision, product launches including Sakana Fugu, Japan’s AI strategy, and his hopes for Japanese society. Watch the interview. -

Vector Search Perf module for slow AI retrieval diagnosis
By
–
2/ Next problem → sluggish AI retrieval. Scaling it up is notoriously difficult. This Vector Search Perf module walks you through diagnosing slow queries with Atlas Metrics. PLUS actually managing memory sizing and quantization in full production > https://
fandf.co/4b6Q3uW -

Unify vector search with Voyage AI retrieval pipelines
By
–
1/ First: aren't you tired of juggling a separate vector DB just to make semantic search work? You can keep it all unified. The Voyage AI course shows exactly how to build advanced two-step retrieval pipelines where your original data already lives. > https://
fandf.co/3SmvV1z -
Seedance 2.5: 30s generations with up to 50 references
By
–
Wow. Seedance is really good at turning greyboxed 3d references into final quality pixels. Like really good.
— Bilawal Sidhu (@bilawalsidhu) 24 juin 2026
Seedance 2.5 coming with 30 second generations and up to 50 (!) references is gonna be mad. Try the workflow Reid is running below. pic.twitter.com/OtbSzfCFhZWow. Seedance is really good at turning grayed-out 3D references into final-quality pixels. Really good. Seedance 2.5 is coming with 30-second generations and up to 50 (!) references — it's going to be insane. Try the workflow Reid runs below.
-

Beginning of a major architectural change in AI development
By
–
Delighted to collaborate with @OpenRouter Products like OpenRouter Fusion and Sakana Fugu have sparked a serious conversation about dependency and resilience in AI. I believe this is only the beginning of a major architectural change to come in the development of
-
AI solves 18 medical cases of children abandoned by doctors
By
–
Urgent: An AI just solved 18 medical cases that doctors had abandoned. All children with rare diseases that no one could identify. Here is how the AI did it:
-
Tips for Automatically Cleaning Duplicate Browser Tabs with Codex
By
–
Here's a little browser use trick I've been loving recently – it's very smooth. I'm the kind of heavy user who opens dozens or hundreds of tabs at once, and they often end up duplicated. Later, I added a trigger task in Codex. Every time I lock my screen or go offline, Codex automatically calls browser use to clean up and close redundant, duplicate tabs – completely hands-free. This trick works with Codex and CC.