Search papers, labs, and topics across Lattice.
This paper introduces CUE Bench, a benchmark designed to enhance emotion understanding in Chinese discourse by focusing on Affective Stance, which encompasses explicit and implicit emotional expressions and their pragmatic intents. The authors argue that existing emotion benchmarks fail to capture the complexity of affective communication, leading to insensitivity in evaluating model performance on nuanced emotional cues. Experimental results demonstrate that integrating Affective Stance significantly improves fine-grained emotion recognition and pragmatic intent detection, highlighting the benchmark's effectiveness in addressing the limitations of current methodologies.
Affective Stance integration boosts emotion recognition accuracy by 3.5% and pragmatic intent detection by 7.8%, revealing hidden layers of emotional communication in discourse.
Emotion understanding in discourse requires reasoning beyond surface sentiment because speakers often convey affect through indirect, implicit, polite, ironic, or deliberately mismatched expressions. Existing emotion benchmarks mainly annotate surface polarity or final emotion categories, while lacking a structured account of how explicit expression, implicit affect, pragmatic intent, and fine grained emotion interact. This limitation makes current evaluations insensitive to cases where affective meaning is concealed, weakened, inverted, or pragmatically reshaped, thereby obscuring model failures in deeper emotion understanding. To address this gap, we introduce CUE Bench, a Chinese Unsaid Emotion benchmark that centers on Affective Stance and covers diverse communicative scenarios. CUE Bench constructs nine human interpretable affective stances from explicit implicit polarity interaction and further provides intent and fine grained emotion annotations for structured affective inference. Experiments show that incorporating Affective Stance improves fine grained emotion recognition by 3.5 percentage points and pragmatic intent detection by 7.8 percentage points over strong baselines.