Search papers, labs, and topics across Lattice.
Affiliation:
5
0
6
Emotion-sensitive neurons in multimodal models reveal shared mechanisms for recognizing emotions across speech and faces, with implications for enhancing emotion recognition capabilities.
Emotion-sensitive neurons in LALMs are language-specific, but pooling cross-lingual evidence reveals powerful, transferable Multilingual Emotion Neurons that enhance affective control.
Collapsed Effective Operators can significantly enhance spectral clustering and neural network architectures by effectively encoding long-range topological interactions in a single operator.
Tired of LLM judges hallucinating when evaluating long, detailed speech captions? EmoSURA offers a more reliable, audio-grounded alternative by verifying atomic perceptual units.
CueNet achieves robust audio-visual speaker extraction under visual degradation by cleverly disentangling and integrating speaker information, acoustic synchronisation, and semantic synchronisation cues, without needing training on degraded visual data.