Search papers, labs, and topics across Lattice.
This paper investigates how speech representations derived from hand-crafted acoustic features and self-supervised learning (SSL) embeddings correlate with cognitive performance at task, domain, and global levels in individuals with mild cognitive impairment (MCI). Analyzing 5,754 German neuropsychological assessments, the study reveals that while SSL embeddings initially outperform hand-crafted features, the trend reverses for MCI classification. The authors identify "specialist" and "generalist" representation behaviors based on task structure, linking task constraints to assessment hierarchy in automated clinical speech analysis.
SSL embeddings, typically superior for speech tasks, surprisingly lose their edge to hand-crafted acoustic features when classifying mild cognitive impairment from speech, challenging assumptions about representation learning in clinical applications.
This study examines the relationship between speech representations and the hierarchical structure of cognitive assessment in mild cognitive impairment. Utilizing 5,754 German neuropsychological assessment recordings, we evaluate six cognitive tasks across three score levels: task, domain, and global levels. We compare hand-crafted acoustic features with self-supervised learning (SSL) embeddings. Results show that although SSL representations generally outperform hand-crafted features at lower levels, this trend reverses for MCI classification. Furthermore, task-specific constraints influence performance: tasks with greater response freedom exhibit performance dilution as hierarchical levels increase, suggesting ``specialist''representations, whereas the performance of highly structured tasks increases toward higher levels, suggesting ``generalist''representations. These findings show links between task constraints and assessment hierarchy in automated clinical speech analysis.