Search papers, labs, and topics across Lattice.
Chinese University of Hong Kong (Shenzhen), China
2
0
2
10
Real-world conversational dynamics significantly challenge target speaker extraction, revealing that even advanced systems struggle with natural overlap and noise.
Ditch slow, multi-step sampling for target speaker extraction: AlphaFlowTSE achieves faster, one-step generation with improved speaker similarity and real-world generalization.