PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 27, 2024Alexandria Engineering Journal16 citationsOpen Access

Deepfake detection based on cross-domain local characteristic analysis with multi-domain transformer

View Full Paper
MAMuhammad Ahmad AminYHYongjian HuCLChang‐Tsun Li

Key Points

Key points are not available for this paper at this time.

Abstract

Deepfake videos present a significant challenge in the current media landscape. While current deepfake detection methods demonstrate satisfactory performance, there is still room for improvement in their ability to generalize and detect unseen scenarios, particularly those involving imperceptible cues. This paper introduces a novel multi-modal deepfake detection model named SpectraVisionFusion Transformer (SVFT), which incorporates spatial and frequency domain statistical artifacts to improve generalization performance. The SVFT framework uses two different backbone encoder models to take advantage of both spatial and frequency domain cues in video sequences, along with a decoder and classifier, for common cross-attention and classification, respectively. The spatial domain branch uses a convolutional transformer-based encoder to analyze facial visual features. In contrast, the frequency domain branch employs a language transformer encoder. Additionally, we introduce a weighted feature embedding fusion mechanism that integrates spectral-based statistical feature embeddings and visual cues to achieve a more comprehensive and balanced spatial-frequency feature representation. By coordinately analyzing these modalities, our model exhibits improved detection and generalization capabilities in unseen scenarios. Our proposed SVFT model achieved 92.57% and 80.63% accuracy in extensive cross-manipulation and dataset evaluation, respectively, while surpassing the performance of traditional and single-domain-based approaches.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Amin et al. (2024) studied this question.

synapsesocial.com/papers/68e77584b6db6435876ea878https://doi.org/10.1016/j.aej.2024.02.035
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 4 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Artificial Intelligence in Digital Media: The Era of Deepfakes2020 · 235 citations
  2. 2Face2Face2018 · 276 citations
  3. 32016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)2016 · 1,543 citations
  4. 42023 11th International Workshop on Biometrics and Forensics (IWBF)2023 · 11 citations