Directo terminado. Casi 3 horas dando un repaso general de la actualidad de la IA (GPT-5, modelos frontera, imágenes, vídeo, agentes, robótica y mucho más). Y una visión compartida de hacia dónde se mueve la Inteligencia Artificial Lo tenéis ya resubido en Youtube!
MULTIMODAL AI
-
Movie-Grade Visual Control: Text and Audio-Driven Animation Technology
By
–
5. Why this is powerful: Movie-grade visual control: Combination of text-controlled global movement + audio-driven micro-expressions – appears lifelike rather than “doll-like.” Long-form videos: Minutes-long stable sequences via frame compression (similar to “hierarchical
-
Key Strengths: Visual Control, Long-Form Output, Non-Human Characters
By
–
4. In summary, the key strengths are: Movie-grade visual control: Audio + Prompt control global and local movements.
Long-form output: Minutes-long videos through hierarchical patchify method.
Non-human control: Also works with animals, animations, or stylized characters. In -
Audio-Driven Lip Sync with Realistic Facial and Body Movements
By
–
3. A still image + short vocal track creates realistic lip movement with subtle head and eye movements.
— Chubby♨️ (@kimmonismus) 9 septembre 2025
This demonstrates finely tuned audio-driven motion combined with text-based control of body posture. pic.twitter.com/diCk06jb0I3. A still image + short vocal track creates realistic lip movement with subtle head and eye movements. This demonstrates finely tuned audio-driven motion combined with text-based control of body posture.
-
Wan2.2-S2V: Stable Video Generation with Synchronized Gestures
By
–
2. A short spoken sentence is reproduced with synchronized gestures and facial expressions—stable, without image drift.
— Chubby♨️ (@kimmonismus) 9 septembre 2025
Wan2.2-S2V thus demonstrates its ability to deliver longer, consistent video output. pic.twitter.com/u4QrECZZtY2. A short spoken sentence is reproduced with synchronized gestures and facial expressions—stable, without image drift. Wan2.2-S2V thus demonstrates its ability to deliver longer, consistent video output.
-
Open-Source Wan2.2-S2V: Audio-Controlled Cinematic Video Generation
By
–
1. Open source, cinematic look, audio-controlled: a novelty. Today, I'm testing Wan2.2-S2V—an open-source speech-to-video model that generates cinematic facial expressions and body motion from audio. This feature makes it possible to accurately match audio to motion using open
-
AI Innovators Lead with NVIDIA Rubin CPX Hardware
By
–
Learn how AI innovators like @cursor_ai , @runwayml , and @magicailabs are leading the charge with NVIDIA Rubin CPX.
-
Creating Surreal Video with Runway AI Generation
By
–
How to create a surreal oner with Runway.
— Runway (@runwayml) 9 septembre 2025
First, shoot your scene. Then cut each moment into a 5 second export. Bring each clip into Runway and use Aleph to transform them however you like – add things, remove things, relight your scene. Once you have all of your shots, bring… pic.twitter.com/JMa86wOVyfHow to create a surreal oner with Runway. First, shoot your scene. Then cut each moment into a 5 second export. Bring each clip into Runway and use Aleph to transform them however you like – add things, remove things, relight your scene. Once you have all of your shots, bring
