UP TO 95% TOKEN REDUCTION WITHOUT CHANGING THE CODE A Netflix engineer just open-sourced Headroom, and it’s one of the smartest ways I’ve seen to cut LLM costs. It wraps Cursor or Claude in a local proxy to compress your payload before it hits the LLM: → Intelligently shrinks
Up to 95% Token Reduction Without Code Changes
By
–
