The results are absolutely wild: • 10x fewer parameters than independent models
• Higher accuracy than single-mask approaches
• Works across vision, speech, even coordinate-based representations At 75% sparsity, RTL beats everything while using only 38K parameters vs 314K
RTL outperforms with 10x fewer parameters
By
–
