Search papers, labs, and topics across Lattice.
Key Laboratory of Multimedia Trusted Perception and Efficient Computing,
1
0
3
1
VideoLLMs can slash 70% of their tokens and still achieve state-of-the-art performance, thanks to a hierarchical pruning strategy that mirrors video structure and LLM information flow.