Search papers, labs, and topics across Lattice.
Shanghai Jiao Tong University
2
0
5
Despite achieving 84.8% accuracy, multimodal models struggle with long-horizon reasoning in electrical circuits, revealing critical gaps in their understanding of physical conventions.
RL fine-tuning LMMs for tool use can collapse structural formats due to strong pretrained tool priors, but a surprisingly simple fix of targeted format rewards and frame-budget randomization can restore stability and boost performance.