Noise residual fingerprints are extracted by a simple yet effective pre-trained Noiseprint++ model, outperforming the state-of-the-art detectors on generalization ability, and the effectiveness of each module is validated by ablation studies.
Abstract
The rapid advancement of generative artificial intelligence (AI) has made synthetic images remarkably realistic, posing security threats such as misinformation and fraud. It is significant to detect the synthetic image in the manner of passive and blind image authentication. Most existing detectors rely on supervised training with large labeled datasets, leading to high costs and degraded performance on unknown generative models. To attenuate such deficiencies, we propose a training-free detection method. Specifically, noise residual fingerprints are first extracted by a simple yet effective pre-trained Noiseprint++ model. Then multi-scale features are further extracted from such residual by a frozen Vision Transformer (ViT), followed by adaptive weighted fusion. Only a few real image samples are used needed to initialize the clustering centers for unsupervised K-Means, distinguishing real and synthetic images without training. Extensive evaluations on four benchmark datasets show that our proposed scheme achieves an average accuracy of 82.2%, outperforming the state-of-the-art detectors on generalization ability. Superior performance is gained on the popular diffusion type of synthetic images, and the effectiveness of each module is validated by ablation studies. Source code will be publicly available at https://github.com/multimediaFor/NoiseCluSID.
Experimental results demonstrate that the proposed approach effectively identifies deepfake images with high accuracy, making it suitable for applications in digital forensics, media verification, and cybersecurity.
J. Kollu, Mortha Pavan, Putta Vardhan et al.· International Journal of Inn...· 0 citations
The paper suggests a hybrid training system and adaptive thresholding to improve the generalization of cross-datasets in image forgery detection and indicates that mixed-domain training is a practical approach that can reduce dataset bias and increase generalization.
Varsha Thakur, Rohit Agarwal· Journal of Intelligent Decis...· 0 citations
This project presents an explainable deep learning framework for identifying real and AI-generated images using the NASNet architecture and achieves high detection accuracy while providing interpretable visual explanations, making it suitable for digital image verification, media authentication, and cybersecurity applications.
Panduga Mounika, Dr.CH. Buchi Reddy· American Journal of AI Cyber...· 0 citations
This work introduces a dual-branch ensemble framework fusing Semantic Deep Learning with Mathematical Forensic Feature Extraction, highlighting the practicality and scalability of mathematical forensics for real-world deployment.
Without noise, image processing and pattern recognition would be a fait accompli! Image denoising remains a fundamental challenge in practical image processing applications. Deep neural networks (DNNs) have set new performance benchmarks over classical techniques by using supervised learning to predict either the clean image or the noise exclusively. This paper brings both paradigms on the same footing and examines them from the minimum meansquared error (MMSE) estimation perspective, which requires that their outputs must satisfy a linear constraint, referred to as the consistency criterion. We show that jointly optimizing the image prediction and noise prediction models by enforcing the consistency criterion bridges the gap between the two paradigms and improves denoising performance across multiple settings. Experiments show up to 0.49 dB Peak Signal-to-Noise Ratio (PSNR) gain on CBSD68, Kodak24, McMaster, Urban100 datasets, and 0.46 dB on SIDD dataset, across diverse architectures like DnCNN, SwinIR, CycleISP, and Restormer, under both synthetic and real-world noise.
Oindri Haldar, Saptarshi Mandal, C. Seelamantula· International Conference on...· 0 citations
RippleNet is proposed, an AI-generated image detection framework based on local differential signals that adaptively identifies forgery-sensitive regions and constructs multi-directional, multi-scale differential representations within local neighborhoods, explicitly characterizing anomalous patterns in neighborhood statistics.
Jiazhen Yang, Ruijin Jin, Junjun Zheng et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.