Search papers, labs, and topics across Lattice.
2
0
4
3
WIDE achieves a remarkable 55.1% performance boost at 50% sparsity, revolutionizing how LLMs can efficiently allocate computation at the token level.
MADA-RL boosts compact model reasoning accuracy by 2% with 16 times fewer trainable parameters, redefining how critics learn from generators.