If you are in Suzhou attending EMNLP 2025, come to my talk tomorrow on grounding multimodal LLMs with world cultural knowledge. Date/time(local): Friday 7th Nov, 10:30-10:45 AM
Room: A301
Section: Multilinguality and language diversity 3 I will be presenting CulturalGround
MULTIMODAL AI
-

CulturalGround: Grounding Multimodal LLMs with World Cultural Knowledge
By
–
-
Generalist’s New Robot Control Model Demonstrates Autonomous Arm Fluidity
By
–
A mi parecer este es de los mejores vídeos que he visto de brazos robóticos operando en tiempo real con 100% de autonomía resolviendo la tarea con esa fluidez y movimientos tan limpios!
— Carlos Santana (@DotCSV) 5 novembre 2025
Esto proviene del nuevo modelo de control de robots de Generalist 🔥pic.twitter.com/UVXk0znGKqIn my opinion, this is one of the best videos I've seen of robotic arms operating in real time with 100% autonomy, solving the task with that fluidity and such clean movements! This comes from Generalist's new robot control model
-
Accidental Video Creation with Grok’s Generative AI
By
–
This video was created by accident. I was going to upload my book cover to @xai @grok Imagine and prompt it to produce some images. But hit the wrong button and mistakenly asked for a video, with no prompt. Result was unexpected and amazing. pic.twitter.com/GfRjOPLO1g
— Bob Gourley – e/acc (@bobgourley) 5 novembre 2025This video was created by accident. I was going to upload my book cover to @xai @grok Imagine and prompt it to produce some images. But hit the wrong button and mistakenly asked for a video, with no prompt. Result was unexpected and amazing.
-

Google Gemini 2.5 Models Now Available on Databricks
By
–
.
@Google
's Gemini 2.5 models are generally available on Databricks! When speed and scale, and multimodal intelligence matter, Gemini models deliver — ideal for multimodal inputs, long-context RAG, summarization, and classification with fast, cost-efficient throughput. With AI -
Text-to-Video: Unified Malleable Multimedia Format Emerges
By
–
Otro ejemplo que demuestra lo que llevo tiempo diciendo: las ideas escritas en texto se convierten en imágenes que realmente podemos convertir en vídeos manipulables como si fueran simulaciones en 3D.
— Carlos Santana (@DotCSV) 5 novembre 2025
El multimedia se está convirtiendo en un único formato general 100% maleable. pic.twitter.com/zYMwjft974Otro ejemplo que demuestra lo que llevo tiempo diciendo: las ideas escritas en texto se convierten en imágenes que realmente podemos convertir en vídeos manipulables como si fueran simulaciones en 3D. El multimedia se está convirtiendo en un único formato general 100% maleable.
-
Real-time video generation with full user control capabilities
By
–
🔴 ¡EDICIÓN de VÍDEO en TIEMPO REAL!
— Carlos Santana (@DotCSV) 5 novembre 2025
Una clara muestra de a dónde vamos en los próximos 1-2 años: A la generación de vídeos en tiempo real donde tendremos el control absoluto!
Y esto aún está en sus fases iniciales, pero wow pic.twitter.com/wpCngRDsJD¡EDICIÓN de VÍDEO en TIEMPO REAL! Una clara muestra de a dónde vamos en los próximos 1-2 años: A la generación de vídeos en tiempo real donde tendremos el control absoluto! Y esto aún está en sus fases iniciales, pero wow
-
Runway Workflows: Multi-Model Generative Pipeline Control
By
–
Bring even more control and customization to your generative pipeline with Workflows. Combine multiple models, modalities and generative steps to get everything done all in one place.
— Runway (@runwayml) 5 novembre 2025
Learn how with today's episode of Runway Academy. pic.twitter.com/DYzgN5a7wSBring even more control and customization to your generative pipeline with Workflows. Combine multiple models, modalities and generative steps to get everything done all in one place. Learn how with today's episode of Runway Academy.
-

IMO-Bench Development: Gemini Math Capabilities Evolution
By
–
A bit of a history of IMO-Bench and our IMO efforts:
a. We started building IMO-Bench around early 2024, which was the precursor of ProofBench (basic). b. IMO-Bench was first mentioned in the Gemini 1.5 paper around May 2024. At that time, Gemini Math-specialized 1.5 Pro scored -
Google développe un nouvel agent d’images pour Stitch
By
–
BREAKING 🚨: Google is working on a new Image Agent for Stitch, likely to be powered by Nano Banana 2.
— 🚨 AI News | TestingCatalog (@testingcatalog) 5 novembre 2025
Integration with AI Studio and potentially with Lovable is coming as well, + a possibility to generate a project brief!
Double bananas 🍌 pic.twitter.com/ouhXyaD64pBREAKING : Google is working on a new Image Agent for Stitch, likely to be powered by Nano Banana 2. Integration with AI Studio and potentially with Lovable is coming as well, + a possibility to generate a project brief! Double bananas
-
Sora Now Available on Android with VPN Access
By
–
Sora disponible desde ayer en Android. Con una VPN ya podéis grabar vuestros propios cameos 👍 pic.twitter.com/KnQzzPy8i2
— Carlos Santana (@DotCSV) 5 novembre 2025Sora disponible desde ayer en Android. Con una VPN ya podéis grabar vuestros propios cameos