Search papers, labs, and topics across Lattice.
2
0
2
2
Unified generation of temporally structured audio achieves unprecedented speaker similarity and cross-turn consistency without task-specific branches.
SketchSong achieves superior song coherence and richness by explicitly planning arrangements before audio generation, outperforming strong post-trained models without extra optimization.