AI Dynamics

Global AI News Aggregator

About

Dynamic Linear Attention for Adaptive Token Compression

"Dynamic Linear Attention" Most long-context linear models compress tokens using fixed blocks or logarithmic schedules, but long texts are not uniform. Stable segments can be summarized, while the

→ View original post on X — @askalphaxiv