Qwen3.7 plus released. Looks good, but why do they compare their models to GPT-5.4 and Opus 4.6? Anyways, multimodal as well
@kimmonismus
-
Open source weights and 1m context win
By
–
Yeah, but opensource and weights and 1m context. So I’d say it’s a win
-

MiniMax M3: 59% SWE-Bench Pro, ahead of GPT-5.5 and Gemini 3.1 Pro
By
–

MiniMax just dropped M3! It hits 59% on SWE-Bench Pro, edging out GPT-5.5 (58.6%) and beating Gemini 3.1 Pro (54.2%). Trails Opus 4.7 on coding, but leads it on autonomous browsing at 83.5% on BrowseComp. First open model to pack frontier coding, a 1M-token context, and native
-
Agents lack session memory despite context window focus, HydraDB proposes new layer
By
–
For two years the whole conversation was about context window size.
— Chubby♨️ (@kimmonismus) 1 juin 2026
Meanwhile the actual problem never moved: agents don't remember anything between sessions. We kept patching it with RAG and manual context injection and calling that memory.
HydraDB is going at the layer… https://t.co/st09X5C45VFor two years the whole conversation was about context window size. Meanwhile the actual problem never moved: agents don't remember anything between sessions. We kept patching it with RAG and manual context injection and calling that memory. HydraDB is going at the layer
-
Synthetic Data as Key to Making Robots Truly Usable
By
–
synthetic data is the key to make robots really usable
-
NVIDIA Launches Cosmos Coalition for Open World Models
By
–
7/ NVIDIA also launched the Cosmos Coalition – with Black Forest Labs, Runway, Skild AI, Agile Robots, LTX and Generalist – to push open world models forward together. The bet: world models are becoming the intelligence layer for robots and AVs, and NVIDIA wants the open
-
Cosmos 3 Nano and Super models on Hugging Face with datasets
By
–
6/ Two sizes, both live on Hugging Face right now: Cosmos 3 Nano (8B) — runs on a single workstation GPU for real-time robotics
Cosmos 3 Super (32B) — datacenter-grade, max quality Plus six open datasets and full post-training scripts on GitHub. -
Cosmos 3 tops open leaderboards for text, video, physics, and robotics
By
–
5/ And it's not a demo. Cosmos 3 tops the open leaderboards: #1 open model on Artificial Analysis for text→image AND image→video
#1 on Physics-IQ for physics accuracy — ahead of Sora 2
Leads PAI-Bench overall, ahead of Veo 3.1
#1 robot policy on RoboArena
Open weights beating -

Cosmos 3: Synthetic data generation for robotics, 60x faster, including rare edge cases.
By
–
4/ The real unlock is data. Robotics has always been bottlenecked by how little real-world training data exists. Cosmos 3 generates physically-accurate synthetic data up to 60x faster – including the rare edge cases you can't safely film: collisions, near-misses, accidents.
Eval -
Cosmos 3: native action generation for VLM, world model, robot
By
–
3/ Inputs and outputs span text, image, video, audio AND action.
— Chubby♨️ (@kimmonismus) 1 juin 2026
That last one is the big deal. Cosmos 3 was trained natively to generate actions, so the same checkpoint can run as a vision-language model, a video world model, or a robot policy. No multi-model orchestration. pic.twitter.com/PoLsK33ytZ3/ Inputs and outputs span text, image, video, audio AND action. That last one is the big deal. Cosmos 3 was trained natively to generate actions, so the same checkpoint can run as a vision-language model, a video world model, or a robot policy. No multi-model orchestration.
