Search papers, labs, and topics across Lattice.
University of the Chinese Academy of Sciences
2
0
4
Allocating prediction capacity dynamically across semantic slots can boost recommendation accuracy while keeping computational costs in check.
Pruning 77.8% of visual tokens without losing performance could revolutionize the efficiency of multimodal large language models.