Search papers, labs, and topics across Lattice.
2
0
3
By intelligently fusing heterogeneous guidance cues, HAFMat achieves unprecedented accuracy in estimating human materials from a single image.
Current VideoQA models falter in understanding complex narratives, but StoryVideoQA and PlotTree redefine how we tackle deep video comprehension.