Search papers, labs, and topics across Lattice.
2
0
3
0
Current video-language models struggle to interpret the nuanced meanings behind social media videos, often missing the implicit narratives that make them humorous or ironic.
Reasoning VLMs falter under semantic distractions, often mistaking irrelevant cues for evidence, which can lead to incorrect answers.