Search papers, labs, and topics across Lattice.
3
52
5
5
Generating missing multi-view data from diverse driving videos boosts closed-loop driving robustness in edge cases by over 30%.
Embodied navigation agents, already struggling, fall apart when faced with the kinds of messy, real-world sensor and instruction corruptions that NavTrust now exposes.
Unlock human-like spatial reasoning in VLMs with VLM-3R, which reconstructs 3D understanding from monocular video using instruction tuning, bypassing the need for external depth sensors.