Search papers, labs, and topics across Lattice.
Tongji University
4
0
5
5
RynnBrain 1.1 not only outperforms all competitors in embodied cognition tasks but also redefines how robots can be trained for complex manipulation through innovative 3D grounding techniques.
Memory-augmented pose estimation can dramatically improve generalization across diverse object instances, outperforming traditional methods by leveraging accumulated geometric knowledge.
TACO achieves state-of-the-art performance in open-vocabulary video recognition by preserving out-of-distribution alignment, challenging the conventional trade-off between generalization and specialization.
Uniformly sampling frames in video LLMs is leaving crucial temporal information on the cutting room floor: GroundVTS selectively attends to the most informative segments, substantially boosting grounding performance.