Search papers, labs, and topics across Lattice.
Zhejiang University
2
0
3
MedUP reveals that integrating visual perception and language understanding in a single token space can dramatically enhance performance in medical vision-language tasks.
A single distilled model can outperform larger heterogeneous teachers by effectively integrating their strengths without interference.