Search papers, labs, and topics across Lattice.
2
0
3
Natural-language instructions can now handle both zero-shot speech synthesis and surgical acoustic editing within a single unified model, operating at 4-step distilled inference speeds without classifier-free guidance.
Aesthetics-guided training enables LeVo 2 to generate songs that not only sound good but also maintain coherence and detail, outperforming existing models.