Search papers, labs, and topics across Lattice.
Tsinghua University, Tencent
1
0
3
CoMem achieves a 7.83x prefill speedup and drastically reduces memory usage while maintaining high performance on long-context tasks, challenging conventional memory management in LLMs.