Search papers, labs, and topics across Lattice.
1
0
2
Training on the entire deployed matrix can unlock up to 81.4% of a model's capacity that was previously unreachable, leading to unprecedented performance gains in low-rank distillation.