/2 MiniMax opens the model weights of Minimax M3 to the public. M3 is the first open-weight model to combine frontier coding, a 1 million token context window, and native multimodal support (images and video) in one package. That combo was locked behind paid APIs until now. The
MULTIMODAL AI
-
Waymo Sensor Fusion Part 1: LiDAR, Radar, Cameras
By
–
(How Waymo Sees — Part 1)
— Satya Mallick (@LearnOpenCV) 23 juin 2026
LiDAR tells a Waymo where everything is. Radar tells it how fast everything's moving. Cameras add color and texture. No single sensor is enough — autonomy lives in the fusion.
Full deep-dive 👇https://t.co/axILFYfFQW pic.twitter.com/bpPFiO7TD5(How Waymo Sees — Part 1)
LiDAR tells a Waymo where everything is. Radar tells it how fast everything's moving. Cameras add color and texture. No single sensor is enough — autonomy lives in the fusion.
Full deep-dive https://
learnopencv.com/3d-lidar-visua
lization/
… -
OpenAI unveils real-time Bidi 1 voice model
By
–
OPENAI 🔥: An upcoming Bidi 1 voice model will be able to translate in real-time!
— 🚨 AI News | TestingCatalog (@testingcatalog) 23 juin 2026
This will unlock a huge pile of use cases to be built on top of when it lands on the APIs. pic.twitter.com/95sRnSzJfsOPENAI: A coming real-time Bidi 1 voice model will enable seamless translation! This will unlock a massive stack of use cases to build upon when it arrives on the APIs.
-
OpenAI’s new bidi voice mode seems crazy
By
–
OpenAI’s new upcoming „bidi“-voice mode sounds insane! pic.twitter.com/9UMzlCgEm9
— Chubby♨️ (@kimmonismus) 23 juin 2026OpenAI's upcoming new voice mode, the "bidi" mode, seems completely crazy!
-
Seedance 2.5 unveiled, Veo 4 missing, Seedance unmatched
By
–
Seedance 2.5 released. It looks insane! Still trying to figure out where Veo 4 is and why nothing comes close to Seedance pic.twitter.com/OlJBRjfwWO
— Chubby♨️ (@kimmonismus) 23 juin 2026Seedance 2.5 is out. It looks crazy! I'm still trying to understand where Veo 4 went and why nothing comes close to Seedance.
-
Seedance 2.5: 4K, 3D models, AI copyright platform from ByteDance
By
–
Seedance 2.5 🤯
— Linus ✦ Ekenstam (@LinusEkenstam) 23 juin 2026
→ 30-second clips
→ Native support for 4K
→ Supports 50 full-modal reference materials, INSANE inc.
→ Supports 3D white models
Simultaneously launches AI copyright commercialization platform.
Bytedance is leading the way here pic.twitter.com/Le6X0m18QLSeedance 2.5 → 30-second clips
→ Native support for 4K
→ Supports 50 full-modal reference materials, INSANE inc. → Supports 3D white models Simultaneously launches AI copyright commercialization platform. Bytedance is leading the way here -
Midjourney’s strange atmospheric cities despite health pivot
By
–
I know they are pivoting to health care(?!) but there is still nothing like Midjourney for making strange and atmospheric images and short animations in ways no other AI image generator can do.
— Ethan Mollick (@emollick) 23 juin 2026
Here are some strange cities I made with similar prompts but very different styles. pic.twitter.com/oGJ74ZxoR8I know they are pivoting to health care(?!) but there is still nothing like Midjourney for making strange and atmospheric images and short animations in ways no other AI image generator can do. Here are some strange cities I made with similar prompts but very different styles.
-

Mastering PyTorch book covering deep learning from CNNs to LLMs
By
–
"Mastering PyTorch: Create and deploy deep learning models from CNNs to multimodal models, LLMs, and beyond" – http://
amzn.to/40IFEQR via @PacktDataML —————
#AI #ML #MachineLearning #DataScience #DataScientist -

VLM³ proves vision-language models are native 3D learners
By
–
What if standard vision-language models already understand 3D—without complex architecture changes or special losses? Meta and Princeton University present VLM³, showing that VLMs are native 3D learners. Their recipe: unify camera focal lengths, use text-based pixel references,


