Search papers, labs, and topics across Lattice.
The University of Texas at Austin
3
0
4
Conditional branching in navigation tasks reveals hidden failures in agent decision-making that standard metrics overlook.
Achieving robust 3D spatial reasoning without any training, ViewMind3D redefines the landscape of 3D question answering.
Captions selected with VEGAS align significantly better with human attention, boosting retrieval performance and challenging the status quo of video captioning metrics.