PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
December 1, 2013101 citationsOpen Access

Deep maxout networks for low-resource speech recognition

YMYajie MiaoFMFlorian MetzeSRShourabh Rawat

Key Points

Key points are not available for this paper at this time.

Abstract

As a feed-forward architecture, the recently proposed maxout networks integrate dropout naturally and show state-of-the-art results on various computer vision datasets. This paper investigates the application of deep maxout networks (DMNs) to large vocabulary continuous speech recognition (LVCSR) tasks. Our focus is on the particular advantage of DMNs under low-resource conditions with limited transcribed speech. We extend DMNs to hybrid and bottleneck feature systems, and explore optimal network structures (number of maxout layers, pooling strategy, etc) for both setups. On the newly released Babel corpus, behaviors of DMNs are extensively studied under different levels of data availability. Experiments show that DMNs improve low-resource speech recognition significantly. Moreover, DMNs introduce sparsity to their hidden activations and thus can act as sparse feature extractors.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Miao et al. (2013) studied this question.

synapsesocial.com/papers/6a20777ca40bcbc79e099a4dhttps://doi.org/10.1109/asru.2013.6707763
Ask AI
Helpful
Bookmark
Share
View Full Paper