What happens when you let a model decide how to allocate FLOPs? Introducing Compute Where It Counts, a new trainable sparsity paradigm released by @crystalAIorg that beats SOTA methods, enabling 80% sparsity and 4x+ speedups on CPU.
Model-Driven FLOP Allocation Achieves 80% Sparsity, 4x CPU Speedups
By
–
