Search papers, labs, and topics across Lattice.
3
0
4
3
Leveraging temporal differences can dramatically enhance video-to-audio generation quality, outperforming even dedicated multimodal representations.
Voice-controlled video generation just got a major upgrade with Vidu S1, achieving real-time performance without visual distortion.
AV-SyncBench reveals that existing benchmarks obscure critical differences in audio-visual model performance by conflating temporal and semantic evaluations.