In this letter, a new set of speech feature parameters based on multirate signal processing and the Teager energy operator is introduced. The speech signal is first divided into nonuniform subbands in mel-scale using a multirate filterbank, then the Teager energies of the subsignals are estimated. Finally, the feature vector is constructed by log-compression and inverse discrete cosine transform (DCT) computation. The new feature parameters have robust speech recognition performance in the presence of car engine noise.
No takes yet. Share an insight, caveat, or question.
Jabloun et al. (1999) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: