Search papers, labs, and topics across Lattice.
4
0
5
0
A single model can now seamlessly generate, understand, and edit 3D content, revolutionizing how we approach multimodal AI tasks.
PAR3D reveals that integrating part-aware representations can dramatically enhance 3D scene understanding, outperforming traditional object-centric models.
A training-free feature adjustment pipeline unlocks the power of Visual Geometry Grounded Transformers for stereo vision, achieving state-of-the-art results on KITTI.
Autonomous vehicles can now leverage the rich semantic understanding of VLMs for safer driving without the computational overhead, thanks to a clever training strategy that distills VLM knowledge into a real-time RL policy.