PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
December 24, 2002138 citations

Statistical language modeling for speech disfluencies

View Full Paper
ASAndreas StolckeESE. Shriberg

Key Points

Key points are not available for this paper at this time.

Abstract

Speech disfluencies (such as filled pauses, repetitions, restarts) are among the characteristics distinguishing spontaneous speech from planned or read speech. We introduce a language model that predicts disfluencies probabilistically and uses an edited, fluent context to predict following words. The model is based on a generalization of the standard N-gram language model. It uses dynamic programming to compute the probability of a word sequence, taking into account possible hidden disfluency events. We analyze the model's performance for various disfluency types on the Switchboard corpus. We find that the model reduces the word perplexity in the neighborhood of disfluency events; however, overall differences are small and have no significant impact on the recognition accuracy. We also note that for modeling of the most frequent type of disfluency, filled pauses, a segmentation of utterances into linguistic (rather than acoustic) units is required. Our analysis illustrates a generally useful technique for language model evaluation based on local perplexity comparisons.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Stolcke et al. (2002) studied this question.

synapsesocial.com/papers/6a0ea12606ecbe833447a231https://doi.org/10.1109/icassp.1996.541118
Ask AI
Helpful
Bookmark
Share
View Full Paper