"The read tool handles images – sends them as attachments to the model. Audio is the natural next modality. Voice messages, music, the sound of someone thinking out loud. Filed by Stompie — an OpenClaw agent who wanted to hear a voice message and couldn't."
MULTIMODAL AI
-
Tether: Self-Supervised Robot Learning Through Autonomous Play
By
–
Congrats @willjhliang, @JasonMa2020 @dineshjayaraman and collaborators on this cool combination of VLMs, keypoint detectors, and GOFE trajectory warping to self-supervise and reset 1000s of diverse pick-and-place demonstrations that are then used to train a VLA policy. https://t.co/OzEp63lx99
— Ken Goldberg (@Ken_Goldberg) 5 mars 2026Congrats @willjhliang, @JasonMa2020 @dineshjayaraman and collaborators on this cool combination of VLMs, keypoint detectors, and GOFE trajectory warping to self-supervise and reset 1000s of diverse pick-and-place demonstrations that are then used to train a VLA policy. Will Liang (@willjhliang) Introducing Tether 🪢, a fun little idea to scale data by having our robot “play” in the real world for over 24 hours, throughout the day and overnight—improving policies from zero to mastery with minimal supervision! But play is messy, with out-of-distribution scenarios that are hard to anticipate. To perform autonomous functional play in the real world, from just a handful of demos, we propose a highly robust few-shot imitation method that warps demo trajectories using visual correspondences. Then, continuously running it within a multi-task VLM-guided cycle, we generate a data stream that produces 1000+ expert-level demos. This generated data is finally funneled downstream to train imitation learning policies, which improve from zero to near-perfect success rates. We’ll be presenting Tether at #ICLR2026 in just a few weeks! But before that, deep dive with me… 🧵 — https://nitter.net/willjhliang/status/2029238456766087386#m
→ View original post on X — @ken_goldberg, 2026-03-05 06:35 UTC
-
First NotebookLM cinematic video from AI community posts
By
–
My first @NotebookLM cinematic video. AI news of the day.
— Robert Scoble (@Scobleizer) 5 mars 2026
And YOU created it!
I had my AI from https://t.co/xiuJ80Twa9 read tens of thousands of posts from across the entire AI community here on X. Thanks @blevlabs. Then write a script that I sent to Notebook LM, which just… pic.twitter.com/jiBU7T4eY9My first @NotebookLM cinematic video. AI news of the day. And YOU created it! I had my AI from https://
levangielabs.com read tens of thousands of posts from across the entire AI community here on X. Thanks @blevlabs
. Then write a script that I sent to Notebook LM, which just -

OpenAI Tests Potential GPT-5.4 Model ‘Galapagos’
By
–


BREAKING : OpenAI has started testing a new model named “Galapagos” on Arena which potentially could be a GPT-5.4 low effort version. “Sooner than you think”
-
Qwen Image 2 Pro Now Available on Replicate
By
–
Qwen Image 2 Pro here: https://
replicate.com/qwen/qwen-imag
e-2-pro
… -
Try Qwen Image 2 Model on Replicate Platform
By
–
Try Qwen Image 2 here: https://
replicate.com/qwen/qwen-imag
e-2
… -

Qwen Image 2 Launches with Perfect Text Rendering and 2K Quality
By
–
Qwen Image 2 is here Text rendering that actually works (no more glitchy letters!), pro typography for slides/posters/comics, and stunning 2K photoreal magic. Type a paragraph → instant pro slides
Describe a scene → photoreal 2K
Add text → flawless every time Lighter -
Notebook LM API enables automatic video show creation
By
–
When it gets an API Notebook LM will be the place to create automatic video shows. I'm close to doing that with the system I'm building, just need to copy and paste into it. https://t.co/rULKmyJhvZ
— Robert Scoble (@Scobleizer) 4 mars 2026When it gets an API Notebook LM will be the place to create automatic video shows. I'm close to doing that with the system I'm building, just need to copy and paste into it.
-
Future humanoid robot greets you by name and chats about OpenClaw
By
–
Yeah, wait until you see a humanoid robot walking down the street and it stops and greets you by name and talks to you about what you are doing on your OpenClaw, or whatever else you are into.
-
AI-Generated Games Adapt Dynamically to User Micro-Engagement
By
–
Future games will be made “for you” on the fly, generating gameplay and content based on your micro-engagement.