Search papers, labs, and topics across Lattice.
Nanjing University
1
0
3
Current omni-modal models can score over 66 points in interactive tasks, but they still falter on visual cues and context retention, revealing a critical gap in their utility as assistants.