Search papers, labs, and topics across Lattice.
1
0
1
3
This paper presents Cross-Lingual F5-TTS 2, a simplified framework for transcript-free cross-lingual voice cloning without forced alignment, and makes the syllable-level speaking rate predictor robust to leading and trailing silence through silence-aware augmentation.