What if unlabeled YouTube videos could teach AI to understand 3D scenes? BIGAI & collaborators built an automated data engine that extracts 3D training data from raw internet videos — no manual labeling needed. Their model achieves strong zero-shot results on 3D detection,
RESEARCH
-
Mouse-Inspired Robot Learns Place Recognition Like a Brain
By
–
How a Mouse-Inspired Robot Learns to Recognize Places Like a Brain
— Ronald van Loon (@Ronald_vanLoon) 8 mai 2026
by @lukas_m_ziegler
#Robotics #Engineering #ArtificialIntelligence #Innovation #Technology pic.twitter.com/hTBZe72TsvHow a Mouse-Inspired Robot Learns to Recognize Places Like a Brain
by @lukas_m_ziegler #Robotics #Engineering #ArtificialIntelligence #Innovation #Technology -
Thousands of Vibe-Coded Apps Expose Corporate and Personal Data Online
By
–
Thousands of Vibe-Coded Apps Expose Corporate and Personal Data on the Open Web | WIRED https://
share.google/jakCcHeK2SW64e
K2i
… #viben #vibecoding #coding #coder #AI #anthropic #artificialintelligence @AlbertoEMachado @Eli_Krumova @postoff25 @Khulood_Almani @anand_narang @NutritiousMind -
Even image models fail as world models; video is harder
By
–
Even the image models are not good enough world models yet, gpt-image-2 and nano banana make pretty obvious world model like mistakes. is way harder, so no hope of that anytime soon. Maybe with 2-3 OOMs it gets good enough, but that doesn't feel like a reasonable thing to
-
Agentic coding use cases by @fchollet
By
–
A few major use cases for agentic coding for me: 1. Adhoc data visualizations. Anytime I have a question that can be answered quantitatively, I generate some code to make a plot. 2. Adhoc data annotation UIs. In ML, "make your own dataset" is often the answer, and that used to
-
Demo at OpenAI event London 2025: model trained on UN translators
By
–
Fun fact, I saw a demo of this at an OpenAI event in London in something like October 2025 – over half a year ago. They told us that they trained the model on UN synchronised translators. Not sure why it took so long to release it https://t.co/lwoaaZy7eB
— Peter Gostev (@petergostev) 7 mai 2026Fun fact, I saw a demo of this at an OpenAI event in London in something like October 2025 – over half a year ago. They told us that they trained the model on UN synchronised translators. Not sure why it took so long to release it
-

AlphaEvolve: Gemini-powered coding agent for advanced algorithms
By
–
I often find it more exciting to read about the practical advantages of AI in real-world applications. Back in 2025, I already had the impression that Google's AlphaEvolve was flying under the radar. AlphaEvolve is a Gemini-powered coding agent for designing advanced algorithms.
-

AI developments finance pros should track from MIT Sloan
By
–
Here are the #AI developments that finance pros should be tracking
by Betsy Vereckey @MITSloan Learn more: https://
bit.ly/4eqJtlz #MachineLearning #ArtificialIntelligence #MI -
OpenAI introduces realtime speech-to-speech models for voice agents
By
–
Big day for developers: new realtime audio models are here in the OpenAI API! 🗣️
— Romain Huet (@romainhuet) 7 mai 2026
Fun one to demo: live translation with GPT-Realtime-Translate, and GPT-Realtime-2, our first speech-to-speech reasoning model for voice agents.
Voice is becoming an interface you can actually ship. https://t.co/w2zES0ot7LBig day for developers: new realtime audio models are here in the OpenAI API! Fun one to demo: live translation with GPT-Realtime-Translate, and GPT-Realtime-2, our first speech-to-speech reasoning model for voice agents. Voice is becoming an interface you can actually ship.
-

DGPO Method Advances Token-Level Credit Assignment in RL
By
–
"DGPO: Distribution Guided Policy Optimization for Fine Grained Credit Assignment" RL for reasoning has a credit assignment problem. One reward basically gets spread across the whole chain-of-thought. This paper, DGPO, fixes that by turning policy deviation into a token-level
