Search papers, labs, and topics across Lattice.
2
0
5
17
MMLDSum-LLM outperforms existing models by significantly enhancing key information coverage and cross-modal consistency in long-document summarization.
VLMs can achieve up to 4.2x faster inference by simply skipping redundant pixels before they even enter the Vision Transformer.