February 28, 2005

Efficient training of large neural networks for language modeling

Puntos clave

Los puntos clave no están disponibles para este artículo en este momento.

Resumen

Recently there has been increasing interest in using neural networks for language modeling. In contrast to the well-known backoff n-gram language models, the neural network approach tries to limit the data sparseness problem by performing the estimation in a continuous space, allowing by this means smooth interpolations. The complexity to train such a model and to calculate one n-gram probability is however several orders of magnitude higher than for the backoff models, making the new approach difficult to use in real applications. In this paper several techniques are presented that allow the use of a neural network language model in a large vocabulary speech recognition system, in particular very, fast lattice rescoring and efficient training of large neural networks on training corpora of over 10 million words. The described approach achieves significant word error reductions with respect to a carefully tuned 4-gram backoff language model in a state of the art conversational speech recognizer for the DARPA rich transcriptions evaluations.

Me gusta

Guardar

Cite This Study

Holger Schwenk (Mon,) studied this question.

synapsesocial.com/papers/6a0dbde4389a567298baa8a0 https://doi.org/https://doi.org/10.1109/ijcnn.2004.1381158

Me gusta

Guardar