Search papers, labs, and topics across Lattice.
6
0
8
2
Achieving speech intelligibility that consistently surpasses ground-truth recordings, Phoenix TTS redefines the boundaries of zero-shot TTS and voice conversion.
MMAC reveals stark differences in audio captioning performance across multiple dimensions, challenging the adequacy of current evaluation methods.
PRR slashes decoding latency by up to 40% in long-context LLMs while maintaining accuracy, revolutionizing the efficiency of dynamic sparse attention.
AI reviewers can be gamed by merely altering how research is presented, achieving significant score increases without changing the underlying science.
SARA achieves a groundbreaking balance between high-fidelity audio and precise linguistic alignment, setting a new standard for zero-shot TTS systems.
WavBench exposes the limitations of current spoken dialogue models in handling real-world conversational nuances like colloquialisms and paralinguistics, despite advances in reasoning capabilities.