Search papers, labs, and topics across Lattice.
2
0
4
1
MLLMs can now achieve over 8% improvement in spatial reasoning for egocentric scenes by leveraging a novel Ego-element Graph for enhanced perception.
Current MLLMs are surprisingly bad at understanding human intent in egocentric videos at a step-by-step level, achieving only 33% accuracy on a new benchmark designed to prevent future-frame leakage.