One image + text + camera trajectory = controllable worlds. All on a single GPU.
— NVIDIA AI (@NVIDIAAI) 19 mai 2026
Our research team just released SANA-WM, a 2.6B open source world model natively trained for 60-second video generation with precise camera control. pic.twitter.com/oXHRCnCRdM
One image + text + camera trajectory = controllable worlds. All on a single GPU. Our research team just released SANA-WM, a 2.6B open source world model natively trained for 60-second video generation with precise camera control.