Search papers, labs, and topics across Lattice.
3
0
3
15
SDM reveals that pretrained image models can be transformed into powerful tools for understanding complex video dynamics, outperforming traditional methods with less supervision.
Video generation models could be the key to unlocking general-purpose vision intelligence, outperforming specialized models with far less training data.
Ditch slow, 2D motion proxies: GMOS directly segments moving objects from RGB video in 3D space and time, achieving state-of-the-art speed and accuracy.