Search papers, labs, and topics across Lattice.
3
0
4
Bridging the gap between surface-centric visual geometry priors and volumetric occupancy prediction, GPOcc++ sets a new standard for 3D scene understanding in dynamic environments.
Segment-level explainable forensics can drastically enhance our ability to detect and interpret localized manipulations in lengthy AI-generated videos.
Achieve state-of-the-art object referring-guided scanpath prediction by fusing VLMs with fixation history and segmentation LoRA, demonstrating the power of perception-enhanced vision-language models.