Search papers, labs, and topics across Lattice.
This study reveals that pretrained self-supervised models exhibit a consistent pattern where fake media samples generate lower-magnitude feature representations compared to real media. By framing deepfake detection as an anomaly detection problem, the authors demonstrate that simple statistical measures of feature magnitude can match the performance of more complex detection methods. Additionally, the findings indicate that the effectiveness of this discriminative signal improves with larger foundation models, linking advancements in representation learning to enhanced detection capabilities.
Fake media consistently generates lower-magnitude representations, allowing simple statistical methods to rival complex detection systems.
Pretrained self-supervised representations have emerged as a core component of current deepfake detection methods, yet it remains unclear which of their properties make real and fake media distinguishable. In this work, we uncover a surprisingly consistent phenomenon: across multiple pretrained models, datasets, and both image and video domains, fake samples systematically produce lower-magnitude representations than their real counterparts. Motivated by this finding, we formulate deepfake detection as an anomaly detection problem and show that simple statistics of feature magnitude achieve competitive performance with far more sophisticated deepfake detection methods. We further investigate the origin of this effect and demonstrate that reduced feature magnitude is primarily associated with semantic shifts introduced by fake content, while low-level generative fingerprints play a comparatively smaller role. Finally, we show that this discriminative signal strengthens as the size of the underlying foundation model grows, suggesting that advances in representation learning naturally translate into stronger zero-shot deepfake detectors.