“FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention” With how Long context LLMs are being bottlenecked by KV cache, because every old token keeps consuming GPU memory even when most of it is irrelevant, this paper turns long context into
RESEARCH
-

Need to Integrate VLAs and Models in Production Robotics
By
–
It's not VLAs versus World Models; production robotics needs both, in addition to model-based methods, all integrated by agentic coding. At ICRA last week, I presented a perspective on the divisions in our field, including
-
Early access to Anthropic’s Mythos-class model changes weekends
By
–
Fable 5 is something to pay attention to. This is another "I had early access to the new Mythos-class Anthropic model and I want to tell you what I thought of it" post. I know, I know, it's annoying. But the way I now spend my weekends has completely changed because of
-
Gary Marcus questions unfalsifiable ‘real deal’ claim about LLMs
By
–
what does this even mean, @dwarkesh_sp, “the real deal”? is it even a falsifiable conjecture? what’s the evidence?
— Gary Marcus (@GaryMarcus) 9 juin 2026
and if you can agree that “If you ask an LLM a question it can't answer, sometimes it will just try to imitate reasoning without doing it”, why acknowledge the… https://t.co/4DtoGvFCCJwhat does this even mean, @dwarkesh_sp
, “the real deal”? is it even a falsifiable conjecture? what’s the evidence? and if you can agree that “If you ask an LLM a question it can't answer, sometimes it will just try to imitate reasoning without doing it”, why acknowledge the -
VQAScore: open-source framework for evaluating text-video/image/3D models
By
–
VQAScore, a simple framework for evaluating text-video/image/3d models is opensource and got many many features. Packaged with leading frontier and commonly used open-source multimodal models and simple/clean eval interface. https://t.co/8xy7SjGEqt pic.twitter.com/wrt6wBTQvQ
— Jean de Nyandwi (@Jeande_d) 9 juin 2026VQAScore, a simple framework for evaluating text-video/image/3d models is opensource and got many many features. Packaged with leading frontier and commonly used open-source multimodal models and simple/clean eval interface.
-

Claude Fable 5, Anthropic’s most powerful model on Poe
By
–
Claude Fable 5 is now available on Poe. The most powerful model from Anthropic to date, designed for long and complex tasks: large-scale code migrations, deep research, and agentic sessions that last hours or days. Cutting-edge.
-

Claude-Fable-5 refuses a third of questions on BullshitBench
By
–
Claude-Fable-5 on BullshitBench: it refused a THIRD of questions (with multiple attempts). Excluding refusals, it performed well, worse than some other Claude models (chart below).
-

AI model predicts fire spread, redirects evacuees to safer exits
By
–
#AI model predicts building fire spread, redirecting evacuees to safer exits in real time
by National Institute of Standards and Technology @TechXplore_com Learn more: https://
bit.ly/4xcBdwo #ArtificialIntelligence #EmergingTech #Innovation #Technology #Tech -
Shower thoughts: autoencoders as good addition after LoRA
By
–
This was purely based on shower-thoughts :D.
As someone called out, autoencoders would have been a good addition after LoRA. -
Ablation studies invisible to user humorous quote
By
–
"ablation studies are not visible to the user" 😛
