At Interrupt, @Clay's Head of AI @jeffbarg shared insights from running 350m GTM agents a month.
— LangChain (@LangChain) 24 juin 2026
✅ Caching can cut LLM costs up to 70%
✅ Bounding tool calls often improves quality, not just cost
✅ Fairness queues matter once you have real multi-tenant load
Worth 12 minutes if… pic.twitter.com/2qbvyct3lx
At Interrupt, @Clay
's Head of AI @jeffbarg shared insights from running 350m GTM agents a month. Caching can cut LLM costs up to 70% Bounding tool calls often improves quality, not just cost Fairness queues matter once you have real multi-tenant load Worth 12 minutes if