Search papers, labs, and topics across Lattice.
3
0
4
3
Achieving state-of-the-art performance in speech synthesis, Qwen-Audio-3.0-TTS excels in multilingual support and robustness against challenging audio conditions.
Room embeddings can now be reliably estimated from reverberant speech with a calibrated uncertainty score, enabling selective prediction from just one utterance.
Forget massive datasets – PilotTTS proves you can achieve state-of-the-art text-to-speech with a lean architecture and smart data engineering on just 200K hours of open-source data.