PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 25, 20225 citations

Hybrid Feature Extraction MFCC and Feature Selection CNN for Speaker Identification Using CNN: A Comparative Study

View Full Paper
AAAya Hasan AbdulqaderSASyed AbdulRahman Al-HaddadSASalah Abdo

Key Points

Key points are not available for this paper at this time.

Abstract

Speaker Identification is known as the technology that enables users to access a device by speaking into the microphone and detecting the present talker among a group of speakers. The deformation of the incoming voice signal by external sounds greatly degrades the quality of the speaker identification systems. Noise could impact the efficiency of speaker recognition as well as cause the system to not function appropriately. Therefore, speech enhancement is critical for improving the performance of speaker recognition systems under challenging conditions. The spectral Subtraction method is among the most common approaches offered for audio enhancement since it is simple to apply and requires minimal computation in signal processing. The suggested system combines three primary modules: background noise reduction, feature extraction, and sound classification. First of all, the speech enhancement approach has been used to remove the additive noise. Secondly, a combined strategy of feature extraction that uses Mel Frequency Cepstral Coefficients (MFCC) as a feature extractor to be integrated with a Mel filter bank as a single package has been used. Furthermore, by using extracted features and a deep learning algorithm, this method can recognize the identity of the speaker. A convolutional neural network (CNN) for speech modeling that demonstrates very positive results in classifying participants was applied. This architecture has been built in a text-independent configuration. The dataset was made containing 50 speakers, each speaker has 20 voice samples. The accuracy and precision parameters were used to check the feasibility of this model. Our successful hybrid approach attained an accuracy and precision of 98.46%.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Abdulqader et al. (2022) studied this question.

synapsesocial.com/papers/6a212e9bc4e60624665ac6f7https://doi.org/10.1109/esmarta56775.2022.9935422
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Deep Learning for Diabetic Retinopathy (DR) Classifier2022 · 1 citations
  2. 2Biometric security system: a rigorous review of unimodal and multimodal biometrics techniques2018 · 26 citations
  3. 3Speaker Verification Using Adapted Gaussian Mixture Models2000 · 4,289 citations
  4. 4Performance Enhancement of Automatic Speech Recognition System using Euclidean Distance Comparison and Artificial Neural Network2018 · 10 citations
  5. 5Deep neural networks for small footprint text-dependent speaker verification2014 · 1,103 citations