It’s over. Google’s new image model Nano Banana just killed Photoshop. Say what you want, it edits anything just by describing it — no tools, no layers, no hassle. 10 insane examples you won’t believe: 1/ Asked Nano-Banana for a straight-on portrait — it nailed it instantly.
MULTIMODAL AI
-
DeepMind, Google NASA, Meta advance AI for wildlife conservation healthcare
By
–
DeepMind’s Perch: AI that listens to wildlife for conservation. Google + NASA: AI doctors for astronauts. Meta’s TRIBE: Wins brain modeling gold.
-

Google Projects for Gemini to Run Research on Uploaded Documents
By
–
BREAKING : Google is working on Projects for Gemini! There, it will be able to run Research tasks across uploaded documents.
-
NextStep-1: Autoregressive Image Generation with Continuous Tokens
By
–
NextStep-1
— AK (@_akhaliq) 15 août 2025
Toward Autoregressive Image Generation with Continuous Tokens at Scale pic.twitter.com/WXY2cukV9dNextStep-1 Toward Autoregressive Image Generation with Continuous Tokens at Scale
-
AI Models Leverage Video Features for Viral Adoption Strategy
By
–
It is interesting to see how much effort is going into making ancillary features of the AI models go viral. Ever since the (organic) Studio Ghibli moment, one focus for Grok & Gemini has been on video as a gateway. A challenge has been whether people have creative video ideas.
-
ChatGPT Updates: New Legacy Models and GPT-5 Thinking Mini Available
By
–
Recapping the updates we’ve made to ChatGPT in the past week: – GPT-4o available under “Legacy models” by default for paid users – Paid users can toggle on “Show additional models” in settings to add legacy models like o3 and GPT-4.1, as well as GPT-5 Thinking mini, to the
-
Humanoid and social robots reshape the future of technology
By
–
In the world of humachinekind, robots and particularly humanoid robots, social robots, and personal robots are becoming increasingly prevalent.
-
NotebookLM Video Overviews: Gemini’s Multimodal AI Analysis
By
–
When the @NotebookLM team was building video overviews, they wanted to combine the best of Gemini's multimodality into one feature. The AI host "sees" your sources, processes the information, and then is able to discuss what is truly unique about them.
— Google AI (@GoogleAI) 14 août 2025
For example, we uploaded… pic.twitter.com/xa9rTSoKBCWhen the @NotebookLM team was building video overviews, they wanted to combine the best of Gemini's multimodality into one feature. The AI host "sees" your sources, processes the information, and then is able to discuss what is truly unique about them. For example, we uploaded
-
MIT CSAIL Explores Current State and Future of AI-Powered Robots
By
–
@CBSMornings recently visited MIT CSAIL to understand the state of AI-powered robots & where we're headed.
— MIT CSAIL (@MIT_CSAIL) 14 août 2025
"The robots we have today are primarily demonstrations," says MIT Prof. & CSAIL Director Daniela Rus, but she adds that "we can choose to do extraordinary things." pic.twitter.com/m4cpwTvYYZ@CBSMornings recently visited MIT CSAIL to understand the state of AI-powered robots & where we're headed. "The robots we have today are primarily demonstrations," says MIT Prof. & CSAIL Director Daniela Rus, but she adds that "we can choose to do extraordinary things."
-
DINOv3: Self-Supervised Learning for Large-Scale Vision Models
By
–
A few highlights of DINOv3 👇
— AI at Meta (@AIatMeta) 14 août 2025
1️⃣SSL enables 1.7B-image, 7B-param training without labels, supporting annotation-scarce scenarios including satellite imagery
2️⃣Produces excellent high-resolution features and state-of-the art performance on dense prediction tasks
3️⃣Diverse… pic.twitter.com/6glMC1S3KWA few highlights of DINOv3 SSL enables 1.7B-image, 7B-param training without labels, supporting annotation-scarce scenarios including satellite imagery
Produces excellent high-resolution features and state-of-the art performance on dense prediction tasks
Diverse
