AI Dynamics

Global AI News Aggregator

About

Google Ships Implicit Caching for Gemini API Cost Savings

We just shipped implicit caching in the Gemini API, automatically enabling a 75% cost savings with the Gemini 2.5 models when your request hits a cache We also lowered the min token required to hit caches to 1K on 2.5 Flash and 2K on 2.5 Pro!

→ View original post on X — @officiallogank