Not having access to native imagegen does hold Fable back somewhat. It is really good at making PNGs, etc, but there are lots of areas (including commercially valuable ones like presentations) where having the ability to have multimodal output would be helpful/token efficient.
MULTIMODAL AI
-
Experimentation with Gemma 4 for creative fashion prompts
By
–
I experimented with using modifications of Gemma 4 to create creative prompts repeatedly. There are still a few quirks, but these are all results from the same simple query: "a dynamic fashion photo of a woman".
-
AI context awareness merges multiple sources into one query
By
–
The bigger unlock is context awareness. You can bring:
→ Multiple tabs
→ Documents
→ Images Into one query. That means the system understands your entire research flow, not isolated searches. This is what Chrome’s AI Mode is aiming for, and it changes how we learn and -
Gemini for Science Could Accelerate Next Big Breakthrough
By
–
Gemini for Science Could Accelerate the Next Big Scientific Breakthrough
— Ronald van Loon (@Ronald_vanLoon) 12 juin 2026
by @GoogleDeepMind#ArtificialIntelligence #MachineLearning #EmergingTech #FutureTech pic.twitter.com/mTaw3V7quUGemini for Science Could Accelerate the Next Big Scientific Breakthrough
by @GoogleDeepMind #ArtificialIntelligence #MachineLearning #EmergingTech #FutureTech -
Decentralized Multi-Embodiment Collaboration using One VLA Policy
By
–
CHORUS
— AK (@_akhaliq) 12 juin 2026
Decentralized Multi-Embodiment Collaboration with One VLA Policy pic.twitter.com/QIyWXsMdiLCHORUS Decentralized Multi-Embodiment Collaboration with One VLA Policy
-

Google Unveils Gemini Omni Flash API
By
–



GOOGLE: Gemini Omni Flash will soon be available via APIs for image-to-video, text-to-video, and video editing!
-
Show, Don’t Tell: Provide References for Better Results
By
–
3. Show, Don't Tell
— Replit ⠕ (@Replit) 11 juin 2026
Give Agent something to reference: screenshots, website links, or files. Want it to match a design? Drop in the mockup. The more it has to look at, the closer it gets to what you actually want. pic.twitter.com/g9CosRTvhO3. Show, Don't Tell Give Agent something to reference: screenshots, website links, or files. Want it to match a design? Drop in the mockup. The more it has to look at, the closer it gets to what you actually want.
-
AI Video: From Generation to Direction with Ray 3.2 for Restaurant
By
–
AI video is quietly shifting from generation to direction.
— SONIA (@S0N_IA_) 11 juin 2026
After testing Ray 3.2 on a restaurant launch concept, that became pretty obvious.
I started with two static images:
one exterior shot,
one interior dining scene.
A few minutes later, I had a cinematic… pic.twitter.com/UbZdhzONvsAI-generated video is quietly moving from generation to direction. After testing Ray 3.2 on a restaurant launch concept, this became quite evident. I started from two static images: an exterior shot and an indoor dining room scene.
-
McConaughey uses ElevenLabs Dubbing v2 to preserve iconic voice in multiple languages
By
–
Matthew @McConaughey is celebrating the big kickoff today. Thanks to ElevenLabs Dubbing v2, he can reach fans around the world in their native language while preserving his iconic voice. pic.twitter.com/SFNoLDNrHa
— ElevenLabs (@ElevenLabs) 11 juin 2026Matthew @McConaughey is celebrating the big kickoff today. Thanks to ElevenLabs Dubbing v2, he can reach fans around the world in their native language while preserving his iconic voice.
-
Using Avatars from Assets for Consistent Characters Across Scenes
By
–
The avatars you create or favorite live in your Assets and can be referenced in any prompt box or dragged directly into a generation.
— ElevenLabs (@ElevenLabs) 11 juin 2026
Use them alongside video models to keep characters consistent across scenes. pic.twitter.com/aTAE6f10yxThe avatars you create or favorite live in your Assets and can be referenced in any prompt box or dragged directly into a generation. Use them alongside video models to keep characters consistent across scenes.