PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 10, 2026Scientific Reports0 citationsOpen Access

Phonetic-DeepKANet: a robust audio spoofing detection framework for English and Arabic

MAMuteb AljasemBowling Green State UniversityHIHafsa IlyasUniversity of Engineering and Technology TaxilaAJAli JavedUniversity of Engineering and Technology Taxila

Key Points

  • This research aims to develop a robust framework for detecting audio spoofing attacks in both English and Arabic languages.
  • Developed a novel framework called Phonetic-DeepKANet (PDK-Net) combining deep and acoustic-phonetic feature extraction.
  • Introduced an Arabic audio spoofing dataset to enhance detection capabilities in underrepresented languages.
  • Evaluated performance on multiple datasets including ASVspoof-2019 and ASVspoof-2021.
  • Achieved a min-tDCF of 0.09 on ASVspoof-2019 LA and 0.14 on ASVspoof-2021 LA, outperforming baseline models.
  • The PDK-Net attained an EER of 8.06% on the Arabic spoofing dataset.
  • Ranked third with an EER of 17.55% on the ASVspoof-2021 DF set among 33 challenge participants.

Abstract

Abstract Audio spoofing attacks, specifically deepfakes, are massively used these days to compromise the security of automatic speaker verification-based systems, leading to data breaches and financial scams. Existing audio spoofing countermeasures are not well-generalized and experience issues when detecting unknown spoofing attacks, including deepfake. Moreover, Arabic audio spoofing detection has been largely neglected, primarily due to the scarcity of Arabic language spoofing datasets. This paper proposes a novel dual-modality approach, Phonetic-DeepKANet (PDK-Net), capable of reliable detection of audio spoofing attacks in English and Arabic. The proposed PDK-Net is comprised of a deep feature extraction module incorporating TransRawNet (TR-Net), an acoustic–phonetic feature extraction module, and a Kolmogorov Arnold Network (KAN) classifier. The deep features from TR-Net are complemented with multi-view acoustic–phonetic representations through concatenation and then classified using KAN. This paper also introduces an Arabic audio spoofing dataset to address the limited availability of such datasets and advance the research in audio spoofing detection for underrepresented languages. The proposed method is evaluated utilizing ASVspoof-2019 LA, 2021 LA, and DF, partial spoof, and our Arabic audio spoofing dataset created in this work. Extensive experimentation on multiple datasets, including voice conversion and text-to-speech synthesized samples, algorithm-wise and cross-corpora evaluation demonstrates the effectiveness and generalizability of our method. We attained the best min-tDCF of 0.09 and 0.14 on the ASVspoof-2019 LA and ASVspoof-2021 LA datasets, respectively, compared to baseline models. However, for the Arabic spoofing dataset, the PDK-Net achieved an EER of 8.06%. It is noteworthy that our method performed best for detecting LA attacks over all 41 methods reported in the ASVspoof-2021 challenge. Further, our method registered the third-best EER of 17.55% amongst 33 challenge participants on the ASVspoof-2021 DF set. These results demonstrate the effectiveness and improved generalization of our approach while detecting unknown spoofing attacks, including codec compressions, channel variations, and encoding artifacts.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Aljasem et al. (2026) studied this question.

synapsesocial.com/papers/6a002126c8f74e3340f9c020https://doi.org/10.1038/s41598-026-51950-9
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Robust DeepFake Audio Detection via an Improved NeXt-TDNN with Multi-Fused Self-Supervised Learning Features2025 · 9 citations
  2. 2ArFake: A Robust Framework for Multi-Dialect Arabic Speech Spoofing Detection Benchmark2025
  3. 3AFAD-MSA: Dataset and Models for Arabic Fake Audio Detection2026 · 1 citations
  4. 4A Survey on Speech Deepfake Detection2024 · 9 citations
  5. 5Fake speech detection using VGGish with attention block2024 · 15 citations