PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 3, 200148 citations

Quantization-based language model compression

View Full Paper
EWE. W. D. WhittakerBRBhiksha Raj

Key Points

Key points are not available for this paper at this time.

Abstract

This paper describes two techniques for reducing the size of statistical back-off gram language models in computer memory. Language model compression is achieved through a combination of quantizing language model probabilities and back-off weights and the pruning of parameters that are determined to be unnecessary after quantization. The recognition performance of the original and compressed language models is evaluated across three different language models and two different recognition tasks. The results show that the language models can be compressed by up to 60% of their original size with no significant loss in recognition performance. Moreover, the techniques that are described provide a principled method with which to compress language models further while minimising degradation in recognition performance.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Whittaker et al. (2001) studied this question.

synapsesocial.com/papers/69b03ca398a0803b6cb32c15https://doi.org/10.21437/eurospeech.2001-8
Ask AI
Helpful
Bookmark
Share
View Full Paper