PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
November 12, 2010Logopedics Phoniatrics Vocology82 citations

On combining information from modulation spectra and mel-frequency cepstral coefficients for automatic detection of pathological voices

View Full Paper
JAJulián D. Arias-LondoñoJGJuan Ignacio Godino-LlorenteMMMaria Markaki

Key Points

Key points are not available for this paper at this time.

Abstract

This work presents a novel approach for the automatic detection of pathological voices based on fusing the information extracted by means of mel-frequency cepstral coefficients (MFCC) and features derived from the modulation spectra (MS). The system proposed uses a two-stepped classification scheme. First, the MFCC and MS features were used to feed two different and independent classifiers; and then the outputs of each classifier were used in a second classification stage. In order to establish the best configuration which provides the highest accuracy in the detection, the fusion of information was carried out employing different classifier combination strategies. The experiments were carried out using two different databases: the one developed by The Massachusetts Eye and Ear Infirmary Voice Laboratory, and a database recorded by the Universidad Politécnica de Madrid. The results show that the combination of MFCC and MS features employing the proposed approach yields an improvement in the detection accuracy, demonstrating that both methods of parameterization are complementary.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Arias-Londoño et al. (2010) studied this question.

synapsesocial.com/papers/69dc241dea70a37eff9553dbhttps://doi.org/10.3109/14015439.2010.528788
Ask AI
Helpful
Bookmark
Share
View Full Paper