I want new benchmarks, for personal assistants, email assistance, customers support
@officiallogank
-
Evals and Benchmarks Don’t Capture What You Think
By
–
There’s also so many layers to most to really understand what they are capturing, it’s pretty interesting. Most evals and benchmarks do not capture what you would initially assume.
-
AI Evals and Benchmarks Capture Model Progress
By
–
I love AI evals and benchmarks, so cool to see interesting ways of capturing model progress
-
Google Advances Gemini API and AI Studio Developer Features
By
–
We are working on: – Deep research in the Gemini API – Making AI Studio more dev centric – Making Gemini app have some of the AI Studio features And more : )
-
Wild Experience Launch Praised for Awesome User Experience
By
–
This is a truly wild experience and an awesome UX for exploring. Congrats to the team on the launch!
-
Gemini Models and App: Past, Present, and Future
By
–
A conversation with @joshwoodward and @tulseedoshi (two of my favorite people) about the past, present, and future of Gemini models + app pic.twitter.com/Bn8lTFMtuJ
— Logan Kilpatrick (@OfficialLoganK) 28 mai 2025A conversation with @joshwoodward and @tulseedoshi (two of my favorite people) about the past, present, and future of Gemini models + app
-
AI Reasoning Toggle Feature Coming June 2026
By
–
The ability to turn off reasoning is coming in June! The team is working hard, and sadly is not as simple as a 2 step prompt : )
-
AI Model Reasoning Now Offers Summarized Thoughts Option
By
–
The model has always reasoned with full thoughts (and continues to do so), now we just have an option to return the summarized thoughts
-
Coding and math conversations provide valuable technical insights
By
–
Some useful info might be there, depends on the conversation though, coding + math I could imagine have useful info