Search papers, labs, and topics across Lattice.
2
0
4
2
UnionSparse achieves up to 3.46x faster low-bit sparse LLM inference on edge GPUs by optimizing metadata handling, challenging the status quo in model efficiency.
Forget slow NTTs: Hermes' hybrid dataflow architecture delivers up to 13.6x faster throughput for homomorphic encryption than GPUs.