Search papers, labs, and topics across Lattice.
1
0
3
2
Jointly applying sparsity, quantization, and low-rank approximations can yield up to 5.66% better accuracy than the best existing methods for LLMs.