Search papers, labs, and topics across Lattice.
6
0
11
7
Open-AoE transforms egocentric video capture into a powerful resource for embodied intelligence, making it easier than ever to train robots with human-like manipulation skills.
Decoupling semantic intent from geometric constraints leads to a 47% reduction in physical hallucinations during human-scene interactions.
DT-Guard outperforms larger models in safety classification while maintaining low-latency performance, proving that reasoning supervision can be efficiently internalized.
Visual-Seeker outperforms proprietary models by actively engaging with visual details, redefining multimodal search capabilities.
Gemini Embedding 2's unified multimodal embeddings beat specialized models across diverse tasks and even generalize zero-shot to niche fields like astronomy and culinary arts.
Forget retargeting: RoboForge's physics-optimized pipeline lets humanoids nail text-guided locomotion with better accuracy and stability.