Skip to content
Preprint

Foundation Models are Implicit Deepfake Detectors

Aug 2026 · 0 citations · 81 references
Computer Science

TL;DR

This work uncovers a surprisingly consistent phenomenon: across multiple pretrained models, datasets, and both image and video domains, fake samples systematically produce lower-magnitude representations than their real counterparts.

Abstract

Pretrained self-supervised representations have emerged as a core component of current deepfake detection methods, yet it remains unclear which of their properties make real and fake media distinguishable. In this work, we uncover a surprisingly consistent phenomenon: across multiple pretrained models, datasets, and both image and video domains, fake samples systematically produce lower-magnitude representations than their real counterparts. Motivated by this finding, we formulate deepfake detection as an anomaly detection problem and show that simple statistics of feature magnitude achieve competitive performance with far more sophisticated deepfake detection methods. We further investigate the origin of this effect and demonstrate that reduced feature magnitude is primarily associated with semantic shifts introduced by fake content, while low-level generative fingerprints play a comparatively smaller role. Finally, we show that this discriminative signal strengthens as the size of the underlying foundation model grows, suggesting that advances in representation learning naturally translate into stronger zero-shot deepfake detectors.

View source

Similar papers

Preprint Aug 2026

Understanding Why Foundation Models Work for Diffusion-Generated Image Detection

This work investigates what cues are exploited by foundation-model-based detectors to distinguish real images from diffusion-generated ones and suggests that foundation-model-based detectors succeed by capturing non-semantic low-to-mid frequency distributional discrepancies between real and diffusion-generated images.

D. Cozzolino, G. Poggi, L. Verdoliva · 0 citations
Preprint Aug 2026

Environment-Invariant Subspace Learning for Generalizable Deepfake Detection

This work proposes an innovative Environment-Invariant Subspace Learning (EISL) framework, which aims to disentangle features into orthogonal forgery-relevant invariant factors and environment-related residual factors via a learnable low-rank projection and designs an Environmental Intervention module that generates diverse and challenging intervention pairs.

Shenghao Chen, Hao Jia, Chen Li et al. · 0 citations
Jul 2026

Explainable Deepfake Detection Challenge

The Explainable Deepfake Detection Challenge at ACM Multimedia 2026 is designed to benchmark this joint capability of classification metrics with semantic similarity, simplicity, and intent-aware grounding metrics that assess whether explanations identify the relevant manipulated entities and supporting visual evidence.

Abhijeet Narang, Kartik Kuckreja, Shreya Ghosh et al. · 1 citation
Jul 2026

Uncertainty-Aware Deepfake Detection via Multi-View Structural Learning

An uncertainty-aware deepfake detection framework that identifies manipulations through inconsistencies across complementary evidence sources by introducing Inter-Branch Disagreement Calibration (IBDC), a disagreement-aware uncertainty modeling mechanism that links predictive uncertainty to conflicts among evidence streams.

Muhammad Umar Farooq, Kutub Uddin, Awais Khan et al. · 1 citation
Open access Aug 2026

Wavelet-based features to improve cross-forgery generalization in deepfake detection

This work explores an approach that integrates wavelet-based frequency analysis with deep learning to enhance deepfake detection, and suggests that wavelet sub-bands expose manipulation cues that are useful for detecting unseen fake classes, but they should not be interpreted as a uniform robustness improvement.

Niccolò Marini, Stefano Berretti, Roberto Caldelli · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.