Search papers, labs, and topics across Lattice.
University of Trento
2
0
4
6
MLLMs often struggle with reasoning due to a failure in dynamic cross-modal coordination, but DyCo-RL fixes this by optimizing attention shifts for better performance.
VLLMs can be made much faster without sacrificing accuracy by intelligently merging redundant tokens across space and time using optimal transport.