On the Generalization of Deep Learning Models in Video Deepfake Detection

被引:6
|
作者
Coccomini, Davide Alessandro [1 ]
Caldelli, Roberto [2 ,3 ]
Falchi, Fabrizio [1 ]
Gennaro, Claudio [1 ]
机构
[1] Ist Sci & Tecnol Informaz, I-56124 Pisa, Italy
[2] Natl Interuniv Consortium Telecommun CNIT, I-50134 Florence, Italy
[3] Univ Mercatorum, Fac Econ, I-00186 Rome, Italy
关键词
deepfake detection; deep learning; computer vision; generalization; IMAGE;
D O I
10.3390/jimaging9050089
中图分类号
TB8 [摄影技术];
学科分类号
0804 ;
摘要
The increasing use of deep learning techniques to manipulate images and videos, commonly referred to as "deepfakes", is making it more challenging to differentiate between real and fake content, while various deepfake detection systems have been developed, they often struggle to detect deepfakes in real-world situations. In particular, these methods are often unable to effectively distinguish images or videos when these are modified using novel techniques which have not been used in the training set. In this study, we carry out an analysis of different deep learning architectures in an attempt to understand which is more capable of better generalizing the concept of deepfake. According to our results, it appears that Convolutional Neural Networks (CNNs) seem to be more capable of storing specific anomalies and thus excel in cases of datasets with a limited number of elements and manipulation methodologies. The Vision Transformer, conversely, is more effective when trained with more varied datasets, achieving more outstanding generalization capabilities than the other methods analysed. Finally, the Swin Transformer appears to be a good alternative for using an attention-based method in a more limited data regime and performs very well in cross-dataset scenarios. All the analysed architectures seem to have a different way to look at deepfakes, but since in a real-world environment the generalization capability is essential, based on the experiments carried out, the attention-based architectures seem to provide superior performances.
引用
收藏
页数:12
相关论文
共 50 条
  • [21] Deepfake Video Detection Based on Spatial, Spectral, and Temporal Inconsistencies Using Multimodal Deep Learning
    Lewis, John K.
    Toubal, Imad Eddine
    Chen, Helen
    Sandesera, Vishal
    Lomnitz, Michael
    Hampel-Arias, Zigfried
    Prasad, Calyam
    Palaniappan, Kannappan
    2020 IEEE APPLIED IMAGERY PATTERN RECOGNITION WORKSHOP (AIPR): TRUSTED COMPUTING, PRIVACY, AND SECURING MULTIMEDIA, 2020,
  • [22] A Performance Enhancement of Deepfake Video Detection through the use of a Hybrid CNN Deep Learning Model
    Ikram, Sumaiya Thaseen
    Priya, V
    Chambial, Shourya
    Sood, Dhruv
    Arulkumar, V
    INTERNATIONAL JOURNAL OF ELECTRICAL AND COMPUTER ENGINEERING SYSTEMS, 2023, 14 (02) : 169 - 178
  • [23] Learning Face Forgery Detection in Unseen Domain with Generalization Deepfake Detector
    Tran, Van-Nhan
    Lee, Suk-Hwan
    Le, Hoanh-Su
    Kim, Bo-Sung
    Kwon, Ki-Ryong
    2023 IEEE INTERNATIONAL CONFERENCE ON CONSUMER ELECTRONICS, ICCE, 2023,
  • [24] A Novel Deep Learning Approach for Deepfake Image Detection
    Raza, Ali
    Munir, Kashif
    Almutairi, Mubarak
    APPLIED SCIENCES-BASEL, 2022, 12 (19):
  • [25] IMPROVING THE GENERALIZATION ABILITY OF DEEPFAKE DETECTION VIA DISENTANGLED REPRESENTATION LEARNING
    Hu, Jiashang
    Wang, Shilin
    Li, Xiaoyong
    2021 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2021, : 3577 - 3581
  • [26] Unsupervised Learning-Based Framework for Deepfake Video Detection
    Zhang, Li
    Qiao, Tong
    Xu, Ming
    Zheng, Ning
    Xie, Shichuang
    IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 25 : 4785 - 4799
  • [27] XAI - Empowered Ensemble Deep Learning for Deepfake Detection
    Kumar, Amidela Anil
    Dheepthi Priyangha, S.J.
    Meghana, P.
    Dheeraj, Muppalla
    Aarthi, R.
    2024 15th International Conference on Computing Communication and Networking Technologies, ICCCNT 2024, 2024,
  • [28] Delving into the Local: Dynamic Inconsistency Learning for DeepFake Video Detection
    Gu, Zhihao
    Chen, Yang
    Yao, Taiping
    Ding, Shouhong
    Li, Jilin
    Ma, Lizhuang
    THIRTY-SIXTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE / THIRTY-FOURTH CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE / THE TWELVETH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2022, : 744 - 752
  • [29] A Survey on Deepfake Video Detection
    Yu, Peipeng
    Xia, Zhihua
    Fei, Jianwei
    Lu, Yujiang
    IET BIOMETRICS, 2021, 10 (06) : 607 - 624
  • [30] A Novel Blockchain-Based Deepfake Detection Method Using Federated and Deep Learning Models
    Heidari, Arash
    Navimipour, Nima Jafari
    Dag, Hasan
    Talebi, Samira
    Unal, Mehmet
    COGNITIVE COMPUTATION, 2024, 16 (03) : 1073 - 1091