Here’s everything that happened this week : — @GoogleMaps released 2 new features, Ask Maps to handle your most complex questions about places and trips and Immersive Navigation for intuitive routes, all with some help from the latest Gemini models — New Gemini features
MULTIMODAL AI
-
Gemini Models Transform Google Maps Navigation
By
–
Learn more about how our Gemini models are transforming Google Maps:
-

Gemini Powers Google Maps with Advanced Multi-Step Reasoning
By
–
We’ve reimagined the way our Gemini models power @GoogleMaps. Here are some use cases you can try (and the advancements that make them possible):
— Google AI (@GoogleAI) 13 mars 2026
“Find a well-lit pickleball court that’s usually less busy on Tuesday nights”
➡️ Maps performs multi-step reasoning across 300M+… pic.twitter.com/gONCAjO2eXWe’ve reimagined the way our Gemini models power @GoogleMaps
. Here are some use cases you can try (and the advancements that make them possible): “Find a well-lit pickleball court that’s usually less busy on Tuesday nights” Maps performs multi-step reasoning across 300M+ -
Weekly AI News Roundup: ChatGPT, Google, Anthropic Updates
By
–
Here’s the AI news from the past week — let me know if I missed anything! – @ChatGPTapp adds interactive visuals for math & science
– @Google rolls our Ask Maps
– Google adds Gemini to Docs, Sheets, Slides & Drive
– Google launches Gemini Embedding 2
– @AnthropicAI adds -
Why Humanoid Robots Struggle With Simple Tasks
By
–
this is a *really* good, very accessible piece by @johnpavlus for @QuantaMagazine about why it's so incredibly hard to make humanoid robots work well and generalize tasks. Why Do Humanoid Robots Still Struggle With the Small Stuff? quantamagazine.org/why-do-hu… via @QuantaMagazine
→ View original post on X — @rachelmetz, 2026-03-13 19:13 UTC
-
Real-time video captioning in browser with LFM2-VL WebGPU
By
–
Real-time video captioning in your browser, powered by LFM2-VL
— Maxime Labonne (@maximelabonne) 13 mars 2026
Another exceptional demo by the great @xenovacom! https://t.co/KRvDEpj138Real-time video captioning in your browser, powered by LFM2-VL Another exceptional demo by the great @xenovacom! Xenova (@xenovacom) Real-time video captioning in your browser with @LiquidAI's LFM2-VL model on WebGPU. Sending every frame to a server was never going to be the answer. Imagine the bandwidth, latency and cost. Local inference. No server costs. Infinitely scalable. This is the way. — https://nitter.net/xenovacom/status/2032504624024854673#m
→ View original post on X — @maximelabonne, 2026-03-13 17:46 UTC
-
SigLIP 2: Google’s Advanced Vision-Language Model Evolution
By
–
SigLIP 2: Advancing Vision-Language Understanding Without Contrastive Bottlenecks
— Satya Mallick (@LearnOpenCV) 13 mars 2026
In this episode of Artificial Intelligence: Papers and Concepts, we explore SigLIP 2, the next evolution of Google’s vision–language model designed to better connect images and text through… pic.twitter.com/7TysLsvoXuSigLIP 2: Advancing Vision-Language Understanding Without Contrastive Bottlenecks In this episode of Artificial Intelligence: Papers and Concepts, we explore SigLIP 2, the next evolution of Google’s vision–language model designed to better connect images and text through
-
AI Video Generation Tool VS2.2 Produces Impressive Creative Output
By
–
This video is a masterpiece! I didn’t expect we will be able to get something like it already with VS2.2! Olga cried when she watched it! https://
youtube.com/watch?v=tsf6J8
halkk
… Thank you The Swack -
Create Runway Characters via API for Real-Time Video Agents
By
–
Learn how to create your own Runway Character via the Runway API. And start bringing real-time video agents directly into your apps, products, websites and experiences.
— Runway (@runwayml) 13 mars 2026
Get started at the link below. pic.twitter.com/UJJFJ1Hz0vLearn how to create your own Runway Character via the Runway API. And start bringing real-time video agents directly into your apps, products, websites and experiences. Get started at the link below.
-

Meta Tests Gemini Models for Search Features
By
–

Meta has been testing Gemini models to powers search features on Meta AI. System prompt response