I'd love more clarity on what model is powering the ChatGPT voice mode I'd love it if that voice model could kickoff background agents using GPT-5 for harder problems, maybe saying "let me think a moment…" And I'd love a general bump to the voice mode model
MULTIMODAL AI
-
NVIDIA GTC Showcases Next Era Physical AI Robotics
By
–
Celebrate #NationalRoboticsWeek with a look back at NVIDIA GTC. From autonomous humanoids mastering complex 3D navigation to highly precise surgical arms and real-time conversational delivery bots, the next era of physical AI has arrived.
— NVIDIA (@nvidia) 10 avril 2026
Which robotic application are you most… pic.twitter.com/OmAd4LgyGtCelebrate #NationalRoboticsWeek with a look back at NVIDIA GTC. From autonomous humanoids mastering complex 3D navigation to highly precise surgical arms and real-time conversational delivery bots, the next era of physical AI has arrived. Which robotic application are you most
-
Long AI-generated video created with Seedance 2.0 x Higgs
By
–
Con esto se acaba de demostrar que el vídeo con IA ya es un medio real
— Nico (@nicos_ai) 10 avril 2026
Un video completo de +10 minutos, sin estudio ni equipo y sin manos deformes. Solo Seedance 2.0 x Higgs
El mismo modelo lo puede usar cualquiera ahora mismo 7 días ilimitado, hasta 70% de descuento https://t.co/1u2oLaiUCLCon esto se acaba de demostrar que el vídeo con IA ya es un medio real Un video completo de +10 minutos, sin estudio ni equipo y sin manos deformes. Solo Seedance 2.0 x Higgs El mismo modelo lo puede usar cualquiera ahora mismo 7 días ilimitado, hasta 70% de descuento
-
Music-2.6 and Music-Cover from MiniMax_AI Now Available on Replicate
By
–
Music-2.6 and Music-Cover from @MiniMax_AI is now live on Replicate!
— Replicate (@replicate) 10 avril 2026
Music-2.6: Generate full-length songs or instrumentals from a text prompt, with optional auto-generated lyrics.
Music-Cover: Reimagine any song in a different style — change voice, instruments, genre, and… pic.twitter.com/rAtammB2PAMusic-2.6 and Music-Cover from @MiniMax_AI is now live on Replicate! Music-2.6: Generate full-length songs or instrumentals from a text prompt, with optional auto-generated lyrics. Music-Cover: Reimagine any song in a different style — change voice, instruments, genre, and
-

MMX-CLI: Multimodal Infrastructure for AI Agents
By
–
Introducing MMX-CLI — our first piece of infrastructure built not for humans, but for Agents. Your Agent can read, think, and write. But ask it to sing, paint, or show you a world it's never seen — and it falls silent. Not because it doesn't understand, but because it has no mouth, no hands, no camera. Today, that changes. MMX-CLI gives every Agent seven new senses — image, video, voice, music, vision, search, conversation — powered by MiniMax's full-modal stack, today's SOTA across mainstream omni-modal models. One command: mmxAgent-native I/O. Zero MCP glue. Runs on your existing Token Plan. Two lines to give your Agent a voice: npx skills add MiniMax-AI/cli -y -g npm install -g mmx-cli Then tell it: "you have mmx commands available." It'll learn the rest. Github → github.com/MiniMax-AI/cli Token Plan: platform.minimax.io/subscrib…
→ View original post on X — @scobleizer, 2026-04-10 16:31 UTC
-
Image Classification App Built with Gemma-4-E4B Vision
By
–
An app built with Gemma-4-E4B that classifies images using the model’s vision capabilities.https://t.co/vv7djYRhbk
— Google AI (@GoogleAI) 10 avril 2026An app built with Gemma-4-E4B that classifies images using the model’s vision capabilities.
-

AI Video Generation Achieves Realistic Dance Movement Synthesis
By
–
These sorts of movements used to be impossible with AI.
— fofr (@fofrAI) 10 avril 2026
Seedance 2:
> A 90s era home video, she is street dancing on a warm city street at dusk in baggy 90s clothes to an early 90s hip-hop track, a group of people are around her cheering her moves, especially when she pulls out… https://t.co/425qNNJyh7 pic.twitter.com/XtXmf34GO4These sorts of movements used to be impossible with AI. Seedance 2:
> A 90s era home video, she is street dancing on a warm city street at dusk in baggy 90s clothes to an early 90s hip-hop track, a group of people are around her cheering her moves, especially when she pulls out -

New research framework proposes training tools instead of retraining AI agents
By
–
Stop retraining your AI agents. Train their tools instead. Most AI agents look great in demos. Then they break in production. A new paper from Stanford and Harvard explains why. It introduces a framework that changes how we think about building agents. The core finding: when
-

Meta to Release Muse Spark API Soon
By
–

Meta is planning to release Muse Spark on the APIs soon. Would be curious also to play with Meta’s 9B model if it will ever come out. Soon
-
OpenAI Voice Mode Runs on Weaker Older Model
By
–
I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model – it feels like the AI that you can talk to should be the smartest AI but it really isn't