AI Dynamics

Global AI News Aggregator

About

Qwen 3.5-Flash: Linear Attention + Sparse MoE Breakthrough

Most companies are scaling models UP to get better performance.
Qwen went the opposite direction. Their 3.5-Flash model uses linear attention + sparse MoE architecture. Translation: You get near-frontier performance without needing a data center to run it.

→ View original post on X — @godofprompt