Search papers, labs, and topics across Lattice.
This paper introduces AUCH-Net, a novel Action Unit-based Consistency-aware Hypergraph Network designed to enhance cross-domain few-shot facial expression recognition (CF-FER) by leveraging action units (AUs) as consistent semantic descriptors. The proposed method includes an AU feature learning (AFL) module and a visual feature learning (VFL) module, both guided by innovative loss functions to ensure consistency across domains. Experimental results demonstrate that AUCH-Net significantly outperforms existing state-of-the-art CF-FER methods, highlighting the importance of modeling AU relationships for improved feature transferability.
Modeling action unit relationships can dramatically enhance facial expression recognition across diverse domains, leading to substantial performance gains.
Recently, cross-domain few-shot facial expression recognition (CF-FER) has received considerable attention. However, the performance of existing CF-FER methods is still unsatisfactory due to inferior transferable feature learning under large domain discrepancy and limited target samples. Fortunately, the action units (AUs), which indicate the movements of different facial muscles, provide consistent conceptual semantics for describing expressions within and across domains. Inspired by this, we propose a novel Action Unit-based Consistency-aware Hypergraph Network (AUCH-Net), which constructs consistency-aware hypergraphs on AUs, for CF-FER. Specifically, AUCH-Net presents a new AU feature learning (AFL) module and a new visual feature learning (VFL) module. The AFL module learns AU features under the guidance of a novel relation consistency loss and an AU regularization loss, while the VFL module learns visual features supervised by a relation consistency loss and a classification loss. By learning consistent AU features, AUCH-Net effectively models the connections between AUs and expression categories. As a result, we can bridge the gap between fine-grained facial variations and high-level expression categories, greatly facilitating the learning of transferable feature representations.Extensive experiments on both in-the-lab and in-the-wild datasets show that our method consistently outperforms several state-of-the-art methods. Our results clearly show that modeling the relationships among AUs holds significant potential for FER under cross-domain few-shot scenarios.