This is a nice paper, well executed! @scott_e_reed had this in mind when developing Gato https://
arxiv.org/abs/2205.06175 — I’m glad to see the idea executed with a humanoid and I’d love to see more work along this direction. Gato stood for General AgenT One. Sadly, we weren’t able to
ROBOTICS
-
Paper praised for executing Gato idea with humanoid; more work desired
By
–
-
Skild Brain AI enables robots to handle unfamiliar environments
By
–
Skild Brain is an AI system developed by Skild AI that helps robots operate in unfamiliar environments.
— Amitav Bhattacharjee (@bamitav) 27 juin 2026
It enables different types of robots to complete real-world tasks without prior training or hard-coded instructions.
By controlling robot movements in real time, it makes… pic.twitter.com/PGsCgd9mYHSkild Brain is an AI system developed by Skild AI that helps robots operate in unfamiliar environments. It enables different types of robots to complete real-world tasks without prior training or hard-coded instructions. By controlling robot movements in real time, it makes
-
Using video to learn control representations, touch important
By
–
I think they also emerge from video. This is what I was the most excited about when helping with projects like Veo. My intent was never to create slop videos, but rather to use video to learn representations for control. I must say however that touch is super important and highly
-
Tesla FSD ‘Nice Guy’ Mode Request for Polite Merging
By
–
I wish there was a Tesla FSD "nice guy" mode where it let people merge in front of you, and didn’t wait till the last minute to merge into backed-up lane
-
Next-token prediction compresses latent structure into understanding
By
–
A model trained for next-token prediction is forced to build compressed representations of latent structure in text. Ilya Sutskever correctly refers to this phenomenon as understanding. Here, a model trained for next-step sensor prediction, with a robot that has proprioception… pic.twitter.com/rHh1nFjJxd
— Nando de Freitas (@NandoDF) 27 juin 2026A model trained for next-token prediction is forced to build compressed representations of latent structure in text. Ilya Sutskever correctly refers to this phenomenon as understanding. Here, a model trained for next-step sensor prediction, with a robot that has proprioception
-
Context-trained decision-making for robots is scoped, not open-ended
By
–
Context-trained decision-making for robots is a very different architecture than a general-purpose model. The clear context framing matters: it means the deployment is scoped, not open-ended.
-
Mouse-Inspired Robot Learns Place Recognition Like a Brain
By
–
How a Mouse-Inspired Robot Learns to Recognize Places Like a Brain
— Ronald van Loon (@Ronald_vanLoon) 27 juin 2026
by @lukas_m_ziegler
#Robotics #Engineering #ArtificialIntelligence #Innovation #Technology pic.twitter.com/Xo3VtZLKTiHow a Mouse-Inspired Robot Learns to Recognize Places Like a Brain
by @lukas_m_ziegler #Robotics #Engineering #ArtificialIntelligence #Innovation #Technology -
Physical AI in factories and vehicles requires enhanced cybersecurity
By
–
As AI moves from the digital world into the physical world, cybersecurity becomes even more important.
— Bernard Marr (@BernardMarr) 26 juin 2026
At Bosch ConnectedWorld 2026 in Berlin, one of the biggest themes was physical AI, where AI is embedded into factories, vehicles, robots and industrial systems. That creates… pic.twitter.com/3GFAZf8ZfmAs AI moves from the digital world into the physical world, cybersecurity becomes even more important. At Bosch ConnectedWorld 2026 in Berlin, one of the biggest themes was physical AI, where AI is embedded into factories, vehicles, robots and industrial systems. That creates
-
AI on the Assembly Line: Instant Potato Counting with Minimal Training
By
–
#AI on the Assembly Line: Instant Potato Counting with Minimal Training
by @IlirAliu_ #ArtificialIntelligence #MachineLearning #ML #MI -

HumanEgo: robot learns skills from human egocentric video
By
–
Your robot could learn a new skill just by watching a few minutes of a human wearing smart glasses! University of Maryland presents HumanEgo: a framework that turns 30 minutes of human egocentric video into a zero-shot robot policy. Instead of needing robot data, it extracts