Search papers, labs, and topics across Lattice.
Affiliation:
3
0
4
Fine-tuning just 0.14% of parameters, ENCORE boosts VLM accuracy by 1.43% through innovative entropy-guided cropping and attention techniques.
Pruning visual tokens based on head alignment can retain nearly all performance while drastically reducing computational costs.
Forget hand-tuning: VisPCO automatically finds optimal visual token pruning configurations in VLMs, outperforming predefined strategies across diverse benchmarks.