Search papers, labs, and topics across Lattice.
Chinese Academy of Sciences, University of Chinese Academy of Sciences
3
0
5
Small models can outperform larger counterparts in task planning by leveraging autonomous experience exploration and hindsight training.
CoT reasoning boosts verbal reasoning but falters in visual tasks, revealing a critical gap in multimodal AI capabilities.
MLLMs can now reason about streaming video with significantly improved accuracy and reduced output length thanks to a novel memory-anchored framework that overlaps watching and thinking.