Search papers, labs, and topics across Lattice.
Northeastern University
3
0
5
FlowCTS-OPD boosts performance metrics by over 3% while solving the temporal supervision mismatch that plagues traditional methods.
SER models, often assumed to generalize well to synthesized speech, actually fail miserably, revealing their reliance on spurious correlations rather than genuine emotional understanding.
SLMs still lag behind omni language models in multi-turn conversational style control, as revealed by the new StyleBench benchmark.