Search papers, labs, and topics across Lattice.
Peking University
5
0
8
EcoFrame achieves a remarkable 13.5x inference speedup while maintaining accuracy, revolutionizing how we approach long-video understanding in VLMs.
EchoCache achieves a remarkable 2.46x speedup in audio-driven video generation without sacrificing quality or coherence.
Squeeze your embodied AI models: DyQ-VLA cuts memory footprint by 70% and speeds up inference by 40% without sacrificing performance, all by dynamically adjusting bit-widths based on real-time kinematic data.
VLA models get a 1.73x speedup with only 5-7% overhead thanks to RAPID, a new edge-cloud collaborative inference framework that smartly handles visual noise and motion continuity.
Achieve global-optimal GEMM mapping for spatial accelerators orders of magnitude faster than existing methods by analytically modeling the mapping space geometrically.