Search papers, labs, and topics across Lattice.
1
0
2
3
WIDE achieves a remarkable 55.1% performance boost at 50% sparsity, revolutionizing how LLMs can efficiently allocate computation at the token level.