Search papers, labs, and topics across Lattice.
2
0
2
3
Cross-modal coupling in generative models can dynamically correct contradictions, leading to superior performance in joint image understanding and generation tasks.
Frozen video diffusion models can effectively serve as competitive encoders for a wide range of tasks, merging generation and understanding seamlessly.