BADAS 2.0: new collision prediction system from Nexar.
It's based on V-JEPA 2!!
JEPA is going to save lives.
MULTIMODAL AI
-
Nexar BADAS 2.0: V-JEPA collision prediction system
By
–
-
Kling AI API Integration for Video and Image Generation
By
–
🚨 Kling AI Skill just dropped.
— Futurepedia – Learn to Leverage AI (@futurepedia_io) 16 avril 2026
You can now call Kling’s core video + 4K image powers directly from agents like Claude Code, Cursor, OpenClaw, Codex, Copilot, Opencode & more – all via natural language.
Think:
▪️Text/image → video + intelligent storyboard
▪️4K image gen + image… pic.twitter.com/ujOwsv6atQKling AI Skill just dropped.
You can now call Kling’s core video + 4K image powers directly from agents like Claude Code, Cursor, OpenClaw, Codex, Copilot, Opencode & more – all via natural language. Think:
Text/image → video + intelligent storyboard
4K image gen + image -
SD2 TTS Audio Consistency Testing With Video Wrappers
By
–
I was testing some TTS audio outputs with SD2 yesterday and found it would always change the audio. If I put the same audio in a video wrapper, it's more likely to stay consistent?
-
Robotic Mouth Recreates Human Speech Through Physical Mechanics
By
–
Most voice AI tries to fake human speech.
— Pascal Bornet (@pascal_bornet) 16 avril 2026
This robotic mouth tries to rebuild it.
That is what makes this so fascinating to me.
Instead of relying on text-to-speech software or digital voice models, it recreates speech by copying the physical mechanics of how we actually talk:… pic.twitter.com/mtWTbAbTu1Most voice AI tries to fake human speech. This robotic mouth tries to rebuild it. That is what makes this so fascinating to me. Instead of relying on text-to-speech software or digital voice models, it recreates speech by copying the physical mechanics of how we actually talk:
-
PAI Generates 4K Videos with Consistent Characters for Hollywood
By
–
utopaistudios PAI can now generate 3-minute 4K videos with consistent characters, worlds, and storylines in one place.
— AI Highlight (@AIHighlight) 16 avril 2026
Hollywood productions are already running it.
That is not a demo, That is a new production pipeline. https://t.co/cDe5qcGsRGutopaistudios PAI can now generate 3-minute 4K videos with consistent characters, worlds, and storylines in one place. Hollywood productions are already running it. That is not a demo, That is a new production pipeline.
-

Google’s Gemini 3.1 Flash TTS Now Available on Poe
By
–
Gemini 3.1 Flash TTS is now available on Poe. A fast, high-quality text-to-speech model from Google. A strong fit for voiceovers, audio content, accessibility features, and adding speech to agent workflows. Try it in Poe app on all platforms and in the Poe API
-
TIPSv2: Enhanced Spatial Awareness in Vision-Language Models
By
–
“TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment” This paper propose a foundational image-text encoder with spatial awareness, as VLMs are usually good at describing an image but much worse at grounding where the concepts live. What they found
-
Toyota Unveils CUE7 Basketball-Playing Humanoid Robot
By
–
Toyota has unveiled CUE7, the latest version of its basketball-playing robot https://
youtu.be/Or323Jrtbao?si
=EUNVlxr-MnsBJe01
… via @YouTube #toyota #basketball #humanoidtech #humanoid #robot #Robotics #AI #TechRevolution #TechInnovation #ArtificialInteligence #PhysicalAI @SpirosMargaris -

GPT-IMAGE-2 Rolling Out: Major Generative AI Release
By
–
🚨 BREAKING : GPT-IMAGE-2 IS ROLLINGOUT! pic.twitter.com/FE4ZXPByhc
— CHOI (@arrakis_ai) 16 avril 2026BREAKING : GPT-IMAGE-2 IS ROLLINGOUT!
-
Google Launches Gemini 3.1 Flash TTS for Developers
By
–
Our most expressive and steerable TTS model yet! Designed to give builders granular control over AI-generated speech, Gemini 3.1 Flash TTS is really fun to play with! Available in preview today – for devs via the Gemini API & @GoogleAIStudio + for enterprises on Vertex AI https://t.co/iMiJJnbiIk
— Demis Hassabis (@demishassabis) 16 avril 2026Our most expressive and steerable TTS model yet! Designed to give builders granular control over AI-generated speech, Gemini 3.1 Flash TTS is really fun to play with! Available in preview today – for devs via the Gemini API & @GoogleAIStudio + for enterprises on Vertex AI