GUI-G^2 Gaussian Reward Modeling for GUI Grounding
MULTIMODAL AI
-
AI Image Generation Creates Stunning Paris Travel Presentation
By
–
The prompt I gave them: "Create a visually rich 8-slide presentation featuring an unforgettable Paris travel itinerary, complete with detailed descriptions and stunning imagery." Let’s just say… I wasn’t ready for this ↓
-
New AI-Only Social Video App Launches Early Access
By
–
Some news: We're building the next big thing — the first-ever AI-only social video app, built on a highly expressive human video model. Over the past few weeks, we’ve been testing it in private beta. Now, we’re opening early access: download the iOS app to join the waitlist, or
-
Continuous Learning and Personalization in AI Systems
By
–
In principle I agree, but in practice we have a small niche use case and this (taste) data won't make it out into the outside world. There'll be millions of people who have different taste requirements and continuous learning for them would be a massive unlock. Yes we can fine
-

Best Models for Consistent Character Generation from Images
By
–
We've put together a guide on what we think are the best models for creating consistent characters from a single reference image (no lora training required). https://
replicate.com/blog/generate-
consistent-characters
… -
Runway API Launches Act-Two Motion Capture for Developers
By
–
Act-Two is now available via the Runway API, allowing you to bring our most advanced motion capture directly into your apps, products, platforms and websites.
— Runway (@runwayml) 21 juillet 2025
Learn more and get started at the link below. https://t.co/p3tEpjNgvtAct-Two is now available via the Runway API, allowing you to bring our most advanced motion capture directly into your apps, products, platforms and websites. Learn more and get started at the link below.
-

Gemini advances multi-step reasoning with novel reinforcement learning techniques
By
–
This year was a major paradigm shift, where we can solve problems end to end in natural language. With novel reinforcement learning techniques, we are able to train an advanced Gemini model on multi-step reasoning proof data, which advances the model's capabilities in terms of
-

Gemini Deep Think Wins Gold at International Mathematical Olympiad
By
–
Very excited to share that an advanced version of Gemini Deep Think is the first to have achieved gold-medal level in the International Mathematical Olympiad! , solving five out of six problems perfectly, as verified by the IMO organizers! It’s been a wild run to lead this
-
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
By
–
CSD-VAR
— AK (@_akhaliq) 21 juillet 2025
Content-Style Decomposition in Visual Autoregressive Models pic.twitter.com/BByVwZq8zKCSD-VAR Content-Style Decomposition in Visual Autoregressive Models
-
AI-powered virtual clothing try-on as a practical application
By
–
The ability to try on clothing with a single selfie is one of the most practical uses of AI I’ve seen this year.
