AI Dynamics

Global AI News Aggregator

About

Still: Amortized KV Cache Compaction in a Single Forward Pass

Instead of deleting tokens or optimizing a new prompt-compressed cache, Still learns to synthesize a compact KV cache in a single forward pass. Thus, a small Perceiver per layer reads the cache.

→ View original post on X — @askalphaxiv