Search papers, labs, and topics across Lattice.
2
0
4
5
VLMs can achieve up to 36.95% higher accuracy in multi-image analytical reasoning tasks, outperforming GPT-4.1 by leveraging synthetic data and a novel reinforcement learning strategy.
Learning the generation order in multimodal tasks can boost performance by over 4%鈥攁 game changer for DLMs.