Search papers, labs, and topics across Lattice.
3
0
5
8
MLLMs can now achieve over 8% improvement in spatial reasoning for egocentric scenes by leveraging a novel Ego-element Graph for enhanced perception.
Robots can now learn complex manipulation tasks directly from human demonstrations using only a pair of smart glasses, achieving zero-shot transfer without specialized hardware.
Current MLLMs are surprisingly bad at understanding human intent in egocentric videos at a step-by-step level, achieving only 33% accuracy on a new benchmark designed to prevent future-frame leakage.