Search papers, labs, and topics across Lattice.
Carnegie Mellon University
3
0
5
23
Downstream performance hinges on SSL pre-training language, not the NAC training language, enabling efficient cross-language model reuse.
Training speech separation models on real-world noisy data doesn't have to mean accepting noisy outputs: this method cuts residual noise in half.
Achieve single-pass alignment of multi-talker speech – a feat previously impossible – by modeling overlaps as shuffles.