Search papers, labs, and topics across Lattice.
This paper introduces the Environment-Invariant Subspace Learning (EISL) framework to enhance cross-distribution generalization in deepfake detection by disentangling forgery-relevant features from environment-related factors. The authors address the challenge of spurious correlations caused by environmental variations, such as lighting and style, which hinder the performance of existing models. Experimental results demonstrate that EISL achieves significant improvements in robustness against unseen forgery types and environmental shifts, outperforming or matching leading detection methods across various settings.
Disentangling forgery cues from environmental noise could revolutionize deepfake detection, leading to models that generalize better across diverse conditions.
Cross-distribution generalization remains a critical bottleneck in deepfake detection. While recent efforts leverage the semantic priors of large-scale visual foundation models (VFMs), a noteworthy yet underexplored challenge remains: the susceptibility of these semantic priors to environmental interference from factors such as lighting and style. Crucially, this interference establishes spurious correlations between forgery cues and environmental patterns that severely limit generalization. To address this fundamental challenge, we propose an innovative Environment-Invariant Subspace Learning (EISL) framework. The core contribution of EISL is that it aims to disentangle features into orthogonal forgery-relevant invariant factors and environment-related residual factors via a learnable low-rank projection. To facilitate robust feature disentanglement, we also design an Environmental Intervention module that generates diverse and challenging intervention pairs, simulating out-of-distribution environmental shifts to guide the model toward discovering truly invariant forgery representations. Experiments across cross-dataset, cross-generator, whole-face synthesis, and corruption settings show consistent gains and competitive or leading performance against strong detectors, demonstrating improved robustness to unseen forgery types and environmental variations. This work provides a new perspective and a valuable exploration for understanding and tackling the generalization barriers of VFMs in deepfake detection.