Search papers, labs, and topics across Lattice.
This study develops a multi-scale temporal framework for EEG-based emotion recognition that addresses the limitations of traditional single-window analysis by decomposing EEG signals into multiple temporal windows of varying durations. By employing a shared attention-based encoder and a dynamic fusion module that assigns weights based on sample-specific characteristics, the framework achieves significant improvements in emotion classification accuracy. The results show a maximum accuracy of 65.22% for binary classification and 45.43% for three-class classification, indicating that dynamic fusion and multi-scale analysis can enhance performance beyond conventional methods.
Dynamic fusion of multi-scale EEG signals boosts emotion recognition accuracy, revealing the nuanced interplay of temporal information in mixed emotions.
Mixed emotions represent a clinically relevant but still underexplored target for automatic emotion recognition. EEG provides millisecond-level access to neural activity, yet most EEG pipelines analyze the signal through a single temporal window, thereby fixing the temporal structure available to the model. This study introduces a multi-scale temporal framework for EEG-based emotion recognition. The EEG waveform is decomposed into windows of one or several durations, processed by a shared attention-based encoder, and integrated through a dynamic fusion module that assigns sample-specific weights across temporal scales. The framework is evaluated under a subject-independent protocol in binary and three-class settings, with the three-class task including the mixed affective category. The best results are 65.22% for the two-class task and 45.43% for the three-class task. Both are obtained with three-scale dynamic-fusion configurations and remain substantially above the full-signal baseline. The best-performing temporal scales differ between the two tasks. Dynamic fusion outperforms concatenation in the highest-scoring two-class configuration and slightly exceeds it in the highest-scoring three-class configuration, although these multi-scale settings require substantially more computation than the full-signal baseline.