Search papers, labs, and topics across Lattice.
2
0
5
High binary QA accuracy in music audio-language models can mask significant biases and errors in instrument grounding, revealing the need for more nuanced evaluation methods.
Shifting from token-level to patch-level modeling in TTS can yield a 1.8x speedup and drastically cut memory usage.