On the Generalization of Deep Learning Models in Video Deepfake Detection

被引:6
|
作者
Coccomini, Davide Alessandro [1 ]
Caldelli, Roberto [2 ,3 ]
Falchi, Fabrizio [1 ]
Gennaro, Claudio [1 ]
机构
[1] Ist Sci & Tecnol Informaz, I-56124 Pisa, Italy
[2] Natl Interuniv Consortium Telecommun CNIT, I-50134 Florence, Italy
[3] Univ Mercatorum, Fac Econ, I-00186 Rome, Italy
关键词
deepfake detection; deep learning; computer vision; generalization; IMAGE;
D O I
10.3390/jimaging9050089
中图分类号
TB8 [摄影技术];
学科分类号
0804 ;
摘要
The increasing use of deep learning techniques to manipulate images and videos, commonly referred to as "deepfakes", is making it more challenging to differentiate between real and fake content, while various deepfake detection systems have been developed, they often struggle to detect deepfakes in real-world situations. In particular, these methods are often unable to effectively distinguish images or videos when these are modified using novel techniques which have not been used in the training set. In this study, we carry out an analysis of different deep learning architectures in an attempt to understand which is more capable of better generalizing the concept of deepfake. According to our results, it appears that Convolutional Neural Networks (CNNs) seem to be more capable of storing specific anomalies and thus excel in cases of datasets with a limited number of elements and manipulation methodologies. The Vision Transformer, conversely, is more effective when trained with more varied datasets, achieving more outstanding generalization capabilities than the other methods analysed. Finally, the Swin Transformer appears to be a good alternative for using an attention-based method in a more limited data regime and performs very well in cross-dataset scenarios. All the analysed architectures seem to have a different way to look at deepfakes, but since in a real-world environment the generalization capability is essential, based on the experiments carried out, the attention-based architectures seem to provide superior performances.
引用
收藏
页数:12
相关论文
共 50 条
  • [1] Deepfake video detection using deep learning algorithms
    Korkmaz, Sahin
    Alkan, Mustafa
    JOURNAL OF POLYTECHNIC-POLITEKNIK DERGISI, 2023, 26 (02): : 855 - 862
  • [2] An efficient deepfake video detection using robust deep learning
    Qadir, Abdul
    Mahum, Rabbia
    El-Meligy, Mohammed A.
    Ragab, Adham E.
    AlSalman, Abdulmalik
    Awais, Muhammad
    HELIYON, 2024, 10 (05)
  • [3] AN EFFICIENT DEEP VIDEO MODEL FOR DEEPFAKE DETECTION
    Sun, Ruipeng
    Zhao, Ziyuan
    Shen, Li
    Zeng, Zeng
    Li, Yuxin
    Veeravalli, Bharadwaj
    Yang Xulei
    2023 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP, 2023, : 351 - 355
  • [4] Texture and artifact decomposition for improving generalization in deep-learning-based deepfake detection
    Gao, Jie
    Micheletto, Marco
    Orru, Giulia
    Concas, Sara
    Feng, Xiaoyi
    Marcialis, Gian Luca
    Roli, Fabio
    ENGINEERING APPLICATIONS OF ARTIFICIAL INTELLIGENCE, 2024, 133
  • [5] An Enhanced Deep Learning-Based DeepFake Video Detection and Classification System
    Awotunde, Joseph Bamidele
    Jimoh, Rasheed Gbenga
    Imoize, Agbotiname Lucky
    Abdulrazaq, Akeem Tayo
    Li, Chun-Ta
    Lee, Cheng-Chi
    ELECTRONICS, 2023, 12 (01)
  • [6] Deepfake Detection through Deep Learning
    Pan, Deng
    Sun, Lixian
    Wang, Rui
    Zhang, Xingjian
    Sinnott, Richard O.
    2020 IEEE/ACM INTERNATIONAL CONFERENCE ON BIG DATA COMPUTING, APPLICATIONS AND TECHNOLOGIES (BDCAT 2020), 2020, : 134 - 143
  • [7] DeepFake Detection Using Deep Learning
    Mansoor, Nazneen
    Iliev, Alexander Iliev
    INTELLIGENT COMPUTING, VOL 3, 2024, 2024, 1018 : 202 - 213
  • [8] Hybrid Deep-Learning Model for Deepfake Detection in Video using Transfer Learning Approach
    Pandey, Raksha
    Kushwaha, Alok Kumar Singh
    NATIONAL ACADEMY SCIENCE LETTERS-INDIA, 2024,
  • [9] Spatiotemporal Inconsistency Learning for DeepFake Video Detection
    Gu, Zhihao
    Chen, Yang
    Yao, Taiping
    Ding, Shouhong
    Li, Jilin
    Huang, Feiyue
    Ma, Lizhuang
    PROCEEDINGS OF THE 29TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2021, 2021, : 3473 - 3481
  • [10] Video Transformer for Deepfake Detection with Incremental Learning
    Khan, Sohail Ahmed
    Dai, Hang
    PROCEEDINGS OF THE 29TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2021, 2021, : 1821 - 1828