Search papers, labs, and topics across Lattice.
2
0
5
6
Achieve temporally stable monocular depth estimation by fusing wheel odometry with a pre-trained depth foundation model, mitigating jitter and failures in dynamic environments.
MMHNet proves you can train a video-to-audio model on short clips and have it generalize to generate coherent audio for videos over 5 minutes long.