Search papers, labs, and topics across Lattice.
Cisco Research
3
0
6
0
FeLiX slashes the time-to-target accuracy in federated learning by over 2X, making models far more responsive to real-time user data.
Sangam slashes latency for diffusion language models by intelligently managing prefill and decode processes, revealing a new paradigm for efficient LLM serving.
Serving both image and video diffusion models on the same hardware? GENSERVE's step-level preemption and dynamic resource allocation can boost your service level agreement (SLA) attainment by up to 44%.