Search papers, labs, and topics across Lattice.
Fondazione Bruno Kessler
3
0
5
Fixed 30-second segmentation emerges as the key to robust long-form speech instruction following, outperforming other methods.
Vulnerabilities in speech models are not just a problem for English; they worsen in other languages and with spoken inputs, revealing a critical oversight in AI safety.
Text prompts might be inflating your SLLM's performance: spoken prompts reveal a significant performance gap, especially in low-resource languages.