2/ You can go beyond fixes.
– Add a party hat
– Change the background
– Adjust lighting + remove reflections in one go Don’t know where to start? Just say: “Make it better”
MULTIMODAL AI
-
Beyond fixes: AI image generation enhancement capabilities
By
–
-
AI Coding Agent with Visual Context Changes Game
By
–
AI coding agent with visual context really changes the game.
-
Google DeepMind Launches Genie 3: Revolutionary Generative AI Model
By
–
Congrats @jparkerholder
, @shlomifruchter
, & the Genie & Veo teams! If you are interested to know more, the latest episode of the @GoogleDeepMind Podcast with the brilliant @FryRsquared has just dropped, and is all about Genie 3 & its incredible potential: -
Genie 3 Advanced Spatial Memory Persistence in Simulations
By
–
Genie 3 has advanced spatial memory, when you make changes in the world, they persist in the simulation even when out of view! https://t.co/ppVSDQ2UKI
— Demis Hassabis (@demishassabis) 22 août 2025Genie 3 has advanced spatial memory, when you make changes in the world, they persist in the simulation even when out of view!
-
Real-time Avatar Control Technology Demonstrates AI Integration
By
–
Control an avatar in real time (in this case a dog on a beach!) https://t.co/JdGSag5DDE
— Demis Hassabis (@demishassabis) 22 août 2025Control an avatar in real time (in this case a dog on a beach!)
-
Mirage 2 Revealed: Major Leap Forward in World Model Technology
By
–
1/ Mirage 2 was revealed. Mirage 1 was released just a month ago!
— Chubby♨️ (@kimmonismus) 21 août 2025
The leap in quality is breathtaking.
"Where Mirage 1 revealed the raw potential of a GTA-style world model, Mirage 2 takes a giant step forward: a general-domain world model that empowers you to create, play,… pic.twitter.com/KsE7QmdDwx1/ Mirage 2 was revealed. Mirage 1 was released just a month ago! The leap in quality is breathtaking. "Where Mirage 1 revealed the raw potential of a GTA-style world model, Mirage 2 takes a giant step forward: a general-domain world model that empowers you to create, play,
-

RotBench: Evaluating MLLMs on Image Rotation Detection
By
–
RotBench Evaluating Multimodal Large Language Models on Identifying Image Rotation
-
Audio-Driven Performance Model Demonstrates Impressive AI Capabilities
By
–
amazing use of our audio-driven performance model! thanks for sharing
-
Runway Aleph transforms video environments and characters with AI
By
–
Runway Aleph can seamlessly alter environments, characters and moods wile still maintaining the original motion of your input video. All you need to do is tell Aleph what you want. pic.twitter.com/1nH9M22d0f
— Runway (@runwayml) 21 août 2025Runway Aleph can seamlessly alter environments, characters and moods wile still maintaining the original motion of your input video. All you need to do is tell Aleph what you want.
-

Google Imagen 4 Released: Three Models Pricing and Performance
By
–
Google’s Imagen 4 is here, but is it worth your time? 🎨✨
— Louis-François Bouchard 🎥🤖 (@Whats_AI) 21 août 2025
What’s new
• Family of 3 models:
– Imagen 4 Fast → $0.02 per image, ~2.7s latency (10x faster than Imagen 3)
– Imagen 4 → $0.04 per image
– Imagen 4 Ultra → $0.06 per image, ~10s latency, higher detail (up to 2K… pic.twitter.com/HkbwHzEBa5Google’s Imagen 4 is here, but is it worth your time? What’s new • Family of 3 models: – Imagen 4 Fast → $0.02 per image, ~2.7s latency (10x faster than Imagen 3) – Imagen 4 → $0.04 per image – Imagen 4 Ultra → $0.06 per image, ~10s latency, higher detail (up to 2K