Search papers, labs, and topics across Lattice.
Shanghai Jiao Tong University
3
0
5
4
COSM achieves a remarkable 2.8x improvement in PIM throughput while keeping CPU performance degradation under 2.0%.
Speculative decoding can be sped up by >2x without sacrificing accuracy by rescuing previously rejected tokens that are semantically valid but lexically different.
Unlock the full potential of your GPU's Tensor Cores for 3D Gaussian Splatting with a GEMM-friendly blending transformation that delivers up to 2x speedups.