Search papers, labs, and topics across Lattice.
5
0
5
10
Fixed 30-second segmentation emerges as the key to robust long-form speech instruction following, outperforming other methods.
Vulnerabilities in speech models are not just a problem for English; they worsen in other languages and with spoken inputs, revealing a critical oversight in AI safety.
Only half of speech translation interactions are rated as usable, revealing critical usability gaps that standard evaluations overlook.
Language diffusion models aren't just generative, they're associative memories that reveal a sharp memorization-to-generalization transition detectable via conditional entropy.
Skip the training: SimulU achieves state-of-the-art simultaneous speech translation by cleverly exploiting pre-trained models, opening the door to truly plug-and-play multilingual communication.