AI Dynamics

Global AI News Aggregator

About

Prompt Caching Reduces Latency by 85% on Long Prompts

With prompt caching, you can reuse a book's worth of context across multiple API requests. This can also reduce latency by up to 85% on long prompts. Use cases include coding assistants, large document processing, and agentic tool use. Get started: https://
docs.anthropic.com/en/docs/build-
with-claude/prompt-caching

→ View original post on X — @anthropicai