Search papers, labs, and topics across Lattice.
5
4
7
14
Unlocking head-level control in Diffusion Transformers enables precise motion transfer without the need for retraining, revolutionizing video generation capabilities.
PanoWorld achieves unprecedented performance in panoramic generation by effectively integrating long-range memory with rotation-equivariant representations.
VAORA reduces hallucinated reasoning in VLMs by aligning visual context with action outcomes, leading to better generalization in unseen tasks.
Gemma 4's unified architecture and reasoning mode enable it to outperform larger models in human-rated tasks while maintaining high efficiency.
PixelEyes achieves precise visual localization by separating reasoning from perception, drastically reducing the redundancy in multi-turn visual searches.