Search papers, labs, and topics across Lattice.
This study introduces phoneme-conditional analysis to quantify the acoustic effects of distinctive phonemes in vocal music across nine languages. By comparing marker syllables to matched controls while controlling for singer, melody, and genre, the authors demonstrate that these phonemic differences can be measured along five acoustic dimensions. The resulting song-level profiles achieve an impressive 85.5% accuracy in classifying the language of unaccompanied vocal performances, highlighting the systematic influence of phonological structure on vocal expression.
Phonological structure leaves measurable traces in vocal music, enabling accurate language classification from unaccompanied singing.
The vocal music of each language carries a distinctive sonic identity, even without instrumental accompaniment. We ask whether these differences are measurable and traceable to specific phonemes. To tackle this question, we introduce phoneme-conditional analysis, which isolates the acoustic effect of typologically distinctive phonemes by comparing marker syllables against matched non-marker controls within the same song, holding singer, melody, and genre constant. Across nine typologically diverse languages and thousands of songs, we measure effects along five acoustic dimensions. Song-level profiles built from these effects identify the language of an unaccompanied vocal at 85.5% balanced accuracy in a nine-way classification with folds grouped by artist; whether the separability arises by accumulation of the phoneme-local effects themselves is left open. Our findings suggest that phonological structure leaves systematic and measurable traces in how each language is sung.