Detection of Diffusion-Generated Images Using CBAMEnhanced Pre-Trained CNNs
Abstract
AI-generated images have amplified the need for effective methods to distinguish between real and synthetic visuals. This underscores the need to develop new approaches to ensure data integrity and combat misinformation. While the existing literature predominantly focuses on Generative Adversarial Networks (GAN)-based synthetic images, researchers have largely overlooked the detection of diffusion-based models. This study fills this gap by demonstrating the potential of convolutional block attention module (CBAM)-enhanced convolutional neural networks (CNNs) for the effective detection of diffusion-based synthetic images. In this study, we use CNNs enhanced with the CBAM to propose a novel approach for detecting synthetic images. The CBAM-enhanced model, trained on the CIFAKE dataset, achieved a remarkable accuracy of 97.38% in detecting synthetic images. We integrate pre-trained CNN architectures, such as ResNet50 and DenseNet121, with a CBAM attention mechanism, which enhances performance by focusing on salient spatial and channel information. This approach presents a model that significantly enhances the detection capabilities for distinguishing fake images. Our findings contribute to the field of deepfake detection by providing a robust solution for automated digital image vetting, with implications for AI ethics, security, and broader societal discourse. The implementation details and source code are available at https://github.com/cmpe-dev/Fake-Detector-with-CBAM.
Keywords
References
- Agarwal, S., & Varshney, L. R. (2019). Limits of deepfake detection: A robust estimation viewpoint. arXiv preprint arXiv:1905.03493. https://doi.org/10.48550/arXiv.1905.03493 google scholar
- Aghasanli, A., Kangin, D., & Angelov, P. (2023). Interpretable-through-prototypes deepfake detection for diffusion models. In Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 467–474. https://doi.org/10.1016/j.cvcv.2023 google scholar
- Altuncu, E., Franqueira, V. N., & Li, S. (2024). Deepfake: Definitions, performance metrics and standards, datasets, and a meta-review. Frontiers in Big Data, 7, 1400024. https://doi.org/10.1016/j.fbd.2024.04.024 google scholar
- Baraheem, S. S., & Nguyen, T. V. (2023). AI vs. AI: Can AI detect AI-generated images? Journal of Imaging, 9(10), 199. google scholar
- Baraheem, S. S., Le, T.-N., & Nguyen, T. V. (2023). Image synthesis: A review of methods, datasets, evaluation metrics, and future outlook. Artificial Intelligence Review, 56(10), 10813–10865. https://doi.org/10.1016/j.artificialin.2023.10813 google scholar
- Bird, J. J., & Lotfi, A. (2024). CIFAKE: Image classification and explainable identification of AI-generated synthetic images. IEEE Access, 12, 15642–15650. google scholar
- Charitidis, P., Kordopatis-Zilos, G., Papadopoulos, S., & Kompatsiaris, I. (2020). Investigating the impact of pre-processing and prediction aggregation on the deepfake detection task. arXiv preprint arXiv:2006.07084. https://doi.org/10.48550/arXiv.2006.07084 google scholar
- Corvi, R., Cozzolino, D., Zingarini, G., Poggi, G., Nagano, K., & Verdoliva, L. (2023). On the detection of synthetic images generated by diffusion models. In ICASSP 2023–IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (pp. 1–5). IEEE. google scholar
Details
Primary Language
English
Subjects
Computer Vision, Image Processing
Journal Section
Research Article
Publication Date
June 30, 2026
Submission Date
June 24, 2025
Acceptance Date
March 18, 2026
Published in Issue
Year 2026 Volume: 10 Number: 1