PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 24, 20240 citationsOpen Access

Shallow Diffusion for Fast Speech Enhancement (Student Abstract)

View Full Paper
LYLei YueHarbin Normal UniversityBCBin ChenHangzhou Dianzi UniversityWTWenxin TaiUniversity of Electronic Science and Technology of China

Key Points

Key points are not available for this paper at this time.

Abstract

Recently, the field of Speech Enhancement has witnessed the success of diffusion-based generative models. However, these diffusion-based methods used to take multiple iterations to generate high-quality samples, leading to high computational costs and inefficiency. In this paper, we propose SDFEN (Shallow Diffusion for Fast spEech eNhancement), a novel approach for addressing the inefficiency problem while enhancing the quality of generated samples by reducing the iterative steps in the reverse process of diffusion method. Specifically, we introduce the shallow diffusion strategy initiating the reverse process with an adaptive time step to accelerate inference. In addition, a dedicated noisy predictor is further proposed to guide the adaptive selection of time step. Experiment results demonstrate the superiority of the proposed SDFEN in effectiveness and efficiency.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Yue et al. (2024) studied this question.

synapsesocial.com/papers/68e72a6ab6db6435876a3cffhttps://doi.org/10.1609/aaai.v38i21.30471
Ask AI
Helpful
Bookmark
Share
View Full Paper