Search papers, labs, and topics across Lattice.
Affiliation:
2
0
3
A single linear layer transforms T2AV models into effective voice-cloning systems, achieving record-breaking speaker similarity while slashing inference time by ~30x.
Unlocking precise control over motion dynamics in video generation could revolutionize how we create and edit video content.