Search papers, labs, and topics across Lattice.
2
0
2
Diffusion teachers and online score networks are no longer bottlenecks for distilling autoregressive video models: directly minimizing sample MMD against reference videos improves VBench scores while unlocking 14B post-training on just eight H200 GPUs.
Single-step diffusion models can outperform their multi-step teachers when distributional matching is reformulated as a Wasserstein gradient flow over Gaussian mixtures.