Search papers, labs, and topics across Lattice.
This study explores the automatic detection of stress in speech during the Trier Social Stress Test (TSST) by analyzing acoustic-prosodic features from recordings of 50 participants. Using a processing pipeline that incorporates speaker diarization and machine learning, the researchers achieved stress detection performance significantly exceeding the mean baseline, while also identifying predictors of physiological and affective stress responses. The results highlight speech as a valuable and non-invasive indicator of stress, with implications for behavioral research and clinical assessment.
Speech analysis can predict stress responses with surprising accuracy, revealing insights into human behavior that traditional methods may overlook.
Automatically detecting stress in speech provides an unobtrusive way to gain insights relevant to behavioral research or clinical assessment. This study investigates the automatic differentiation between a stressful and non-stressful situation, and the prediction of physiological and affective stress responses. Speech data was collected from 50 participants who either completed the Trier Social Stress Test (TSST) or a non-stressful control condition. With a processing pipeline that included speaker diarization and machine learning models, we achieved stress detection performance significantly above a mean baseline. Moreover, relevant physiological and affective stress responses were partially predictable from acoustic-prosodic features. Feature-importance analyses identified the most informative predictors contributing to model performance. The findings demonstrate that speech can serve as a meaningful and unobtrusive indicator of multiple dimensions of the human stress response.