Search papers, labs, and topics across Lattice.
3
52
4
4
Generating missing multi-view data from diverse driving videos boosts closed-loop driving robustness in edge cases by over 30%.
Achieving photorealistic 3D human avatars from a single image in under a second could revolutionize virtual reality and gaming applications.
Unlock human-like spatial reasoning in VLMs with VLM-3R, which reconstructs 3D understanding from monocular video using instruction tuning, bypassing the need for external depth sensors.