Nemotron 3 Nano Omni is coming to our Nemotron Labs livestreams. May 5 – How the Developer Community Builds Sub-Agents with NVIDIA Nemotron 3 Nano Omni https://
addevent.com/event/rwnjn543
s4jq
… May 12 – Ask the Experts: Nemotron 3 Nano Omni – https://
addevent.com/event/3myyxpbk
1hgw
… Save the dates and see
MULTIMODAL AI
-
NVIDIA Nemotron 3 Nano Omni Developer Livestreams Announced
By
–
-

20 Must-Try AI Tools for 2025: Complete Stack
By
–
20 must-try AI tools for 2025. This stack covers everything: chat & reasoning image gen video gen voice/music AI search automation AI agents My top 3 right now: ChatGPT, Perplexity, Midjourney. What are yours? #AI #AITools #GenerativeAI #Automation
-

Meta Introduces Tuna-2: Pixel Embeddings for Multimodal AI
By
–
Meta presents Tuna-2 Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation paper: https://
huggingface.co/papers/2604.24
763
… -
Nemotron 3 Nano Omni: NVIDIA’s Fully Open Source AI Model
By
–
Built on NVIDIA’s open ecosystem, Nemotron 3 Nano Omni is fully open source, including:
• Open weights
• Open data
• Open recipes Read the blog for more details -

Nemotron 3 Nano Omni: Unified Multimodal Architecture for AI Subagents
By
–
Nemotron 3 Nano Omni was designed for powering subagents. Instead of stitching together separate models for language, vision, and speech, it ties them into a single architecture that more efficiently feeds context to orchestrators.
-

Nemotron 3 Nano Omni: Efficient Open Multimodal Model Released
By
–
Meet Nemotron 3 Nano Omni 👋
— NVIDIA AI (@NVIDIAAI) 28 avril 2026
Our latest addition to the Nemotron family is the highest efficiency, open multimodal model with leading accuracy.
30B parameters. 256K context length. 🧵👇 pic.twitter.com/j4SPpU9SaIMeet Nemotron 3 Nano Omni Our latest addition to the Nemotron family is the highest efficiency, open multimodal model with leading accuracy. 30B parameters. 256K context length.
-
Microsoft Presents World-R1: 3D Constraints for Video Generation
By
–
Microsoft presents World-R1
— AK (@_akhaliq) 28 avril 2026
Reinforcing 3D Constraints for Text-to-Video Generation
paper: https://t.co/07BWqkSdwd pic.twitter.com/WuRwRZfAmOMicrosoft presents World-R1 Reinforcing 3D Constraints for Text-to-Video Generation paper: https://
huggingface.co/papers/2604.24
764
… -
ChatGPT’s New Image Model Analysis Examples
By
–
Here you have a few more examples of use in my analysis of ChatGPT's new image model
-
OpenAI adds 360° image viewer to ChatGPT
By
–
Mola! Han estado rápido los de OpenAI, que tras haber visto como la gente usaba el modelo de imágenes para crear imágenes 360º, han implementado un visor para poder visualizarlas directamente en ChatGPT
— Carlos Santana (@DotCSV) 28 avril 2026
Es el Genie 3 de aliexpress, pero aún así mola! pic.twitter.com/UgmgumDsBnCool! OpenAI folks have been quick on the draw—after seeing how people were using the image model to create 360º images, they've implemented a viewer so you can check them out right in ChatGPT. It's the Genie 3 from AliExpress, but it still rocks!
-
Perplexity AI merges Google Earth Flight Simulator with GTA gameplay
By
–
A Google Earth + Flight Simulator, fully cooked by Perplexity Computer with Codex/CC as subagents inside it. GTA vibes. pic.twitter.com/TioIvaVpoL
— Aravind Srinivas (@AravSrinivas) 28 avril 2026A Google Earth + Flight Simulator, fully cooked by Perplexity Computer with Codex/CC as subagents inside it. GTA vibes.