Search papers, labs, and topics across Lattice.
South Dakota State University
1
0
2
1
CLIP-CC-Bench reveals that current video-language models struggle with generating coherent long-form descriptions, highlighting a critical gap in their capabilities.