Search papers, labs, and topics across Lattice.
1Zhejiang University
2
0
4
Achieving over 90% performance retention with a staggering 20x KV cache compression could redefine efficiency in long-context audio inference.
Audio LLMs can now be systematically evaluated for character alignment in role-playing scenarios, thanks to a new framework that judges both text and vocal features.