Search papers, labs, and topics across Lattice.
3
0
6
4
Fine-grained cross-modal alignment in audio-video generation can dramatically enhance synchronization and quality, as shown by OmniVAE's innovative training approach.
Action chunk utilization triples and physical execution steps drop by over 50%, resulting in a 5.83x speedup in VLA model deployment without sacrificing performance.
Generative video models can now simulate a continuously evolving world, even when objects are out of sight, thanks to a new framework that maintains persistent global state.