CausalCine Real-Time Autoregressive Generation for Multi-Shot Narratives
MULTIMODAL AI
-
Runway Launches AI Agent for Creative Video Ideation and Editing
By
–
Meet Runway Agent. Your new AI creative partner that helps you ideate and execute fully finished, sound designed and edited videos. All with just a simple conversation. From ads to shorts to content for social, Runway Agent makes it easy to make more of what you need.
— Runway (@runwayml) 13 mai 2026
Get… pic.twitter.com/lFuN6uT6OeMeet Runway Agent. Your new AI creative partner that helps you ideate and execute fully finished, sound designed and edited videos. All with just a simple conversation. From ads to shorts to content for social, Runway Agent makes it easy to make more of what you need. Get
-

OpenAI teases ultrafast mode and image model update for Thursday
By
–
what the heck, openai is cooking – ultrafast mode incoming probably this thursday – + an update to the new image model thats already freaking good openai has such a run lately, love it
-
AI video creation breakdown by PJaccetturo on Gossip_Goblin
By
–
If you want to understand what it takes to make great AI video, this breakdown from @PJaccetturo on the work of Zach @Gossip_Goblin is bringing fire https://t.co/8hAMU7BVkx
— Linus ✦ Ekenstam (@LinusEkenstam) 13 mai 2026If you want to understand what it takes to make great AI video, this breakdown from @PJaccetturo on the work of Zach @Gossip_Goblin is bringing fire
-

EgoMemReason: A Memory-Driven Benchmark for Egocentric Video Understanding
By
–
EgoMemReason A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Understanding
-

SenseNova-U1: Unifying Multimodal Understanding and Generation
By
–
SenseNova-U1 Unifying Multimodal Understanding and Generation with NEO-unify Architecture
-
Identifying visual artifacts in ChatGPT image generation
By
–
¿Habéis detectado ya cuál es el "tinte amarillo" de la nueva versión de imágenes de ChatGPT? Yo lo tengo claro Os dejo un examen visual, pero cuidado porque una vez lo veais no parareis de notarlo.
-
Using Audio Tags to Guide AI Speech Expressiveness
By
–
The old limitation was not clarity. AI voices were already clean and understandable. The real limitation was control. Most AI speech sounded like it was reading a script. Now, with audio tags, we can guide how the AI speaks, not just what it says. That changes the creative
-
Google Gemini 3.1 Flash enables expressive AI voice performance
By
–
AI voice is no longer just text-to-speech.
— Ronald van Loon (@Ronald_vanLoon) 13 mai 2026
It is becoming text-to-performance.
With @Google’s Gemini 3.1 Flash TTS, you can now direct an AI voice almost like an actor:
Tone.
Emotion.
Pacing.
Delivery shifts mid-sentence.
Here’s why that matters… pic.twitter.com/khSiU8knkfAI voice is no longer just text-to-speech. It is becoming text-to-performance. With @Google
’s Gemini 3.1 Flash TTS, you can now direct an AI voice almost like an actor: Tone.
Emotion.
Pacing.
Delivery shifts mid-sentence. Here’s why that matters… -

ELF: Embedded Language Flows — arXiv paper
By
–
ELF: Embedded Language Flows Hu et al.: https://
arxiv.org/abs/2605.10938 #ArtificialIntelligence #AIAgents
