PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 19992,427 citations

Learning to forget: continual prediction with LSTM

View Full Paper
FGFelix A. Gers

Key Points

Key points are not available for this paper at this time.

Abstract

Long Short-Term Memory (LSTM, Hochreiter Schmidhuber, 1997) can solve numerous tasks not solvable by previous learning algorithms for recurrent neural networks (RNNs). We identify a weakness of LSTM networks processing continual input streams that are not a priori segmented into subsequences with explicitly marked ends at which the networks internal state could be reset. Without resets, the state may grow indefinitely and eventually cause the network to break down. Our remedy is a novel, adaptive forget gate that enables an LSTM cell to learn to reset itself at appropriate times, thus releasing internal resources. We review illustrative benchmark problems on which standard LSTM outperforms other RNN algorithms. All algorithms (including LSTM) fail to solve continual versions of these problems. LSTM with forget gates, however, easily solves them in an elegant way. 1 Introduction Recurrent neural networks (RNNs) constitute a very powerful class of computational models, capable of ...

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Felix A. Gers (1999) studied this question.

synapsesocial.com/papers/6a0ec392c12540356222a932https://doi.org/10.1049/cp:19991218
Ask AI
Helpful
Bookmark
Share
View Full Paper