We just shipped implicit caching in the Gemini API, automatically enabling a 75% cost savings with the Gemini 2.5 models when your request hits a cache We also lowered the min token required to hit caches to 1K on 2.5 Flash and 2K on 2.5 Pro!
Google Ships Implicit Caching for Gemini API Cost Savings
By
–