Search papers, labs, and topics across Lattice.
Fudan University
2
0
4
Predicting future actions in videos using foresight expressions reveals a new frontier for spatio-temporal reasoning in video segmentation.
Interactive world models still have a long way to go: a comprehensive benchmark reveals that even state-of-the-art models struggle to consistently perform well across video quality, interaction adherence, and physics compliance.