Search papers, labs, and topics across Lattice.
WeChat HPC, Tencent Inc.
4
0
8
4
Achieving a 4.41x speedup in attention computation while maintaining video quality could redefine efficiency standards in video generation models.
Physically aligned video models can boost robotic manipulation success rates by over 50% compared to traditional methods.
Egocentric human video can outperform traditional teleoperated robot data, achieving superior performance in embodied model pretraining with lower costs and greater diversity.
Current Vision-Language Models get stuck in Pokemon Legends: Z-A not because they can't plan, but because they can't get unstuck.