PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
June 4, 2026Expert Systems0 citations

Research on Memory Algorithm Based on Time Series Model and Reinforcement Learning

View Full Paper
QZQing ZhaoLSLong ShaoYQYaxiu Qiao

Key Points

  • The aim is to integrate time-series modeling with reinforcement learning to enhance memorization methods based on spaced repetition.
  • Developed the GLD-HLR model utilizing Discrete Cosine Transform and Legendre Projection Unit for memory representation.
  • Created a PPO-MMC algorithm to optimize review intervals in a continuous state space and achieve joint learning of memory prediction and policy scheduling.
  • Validated the model's effectiveness through ablation and comparative experiments.
  • Achieved a mean absolute error (MAE) of recall probability predictions below 0.03, a 4% reduction over the LSTM-HLR model.
  • Mean absolute percentage error (MAPE) for half-life predictions was below 0.2, surpassing the prediction accuracy of other models.
  • The PPO-MMC algorithm facilitated learning over 8000 words in 1000 days, with over 7000 memorized at the target half-life.

Abstract

ABSTRACT Spaced repetition is a highly effective method of memorization that helps learners to remember large amounts of content efficiently. This paper presents a spaced repetition framework integrating time‐series modelling with reinforcement learning. We propose the GLD‐HLR model, which utilizes Discrete Cosine Transform (DCT) to decouple multi‐scale temporal features in the frequency domain and a Legendre Projection Unit (LPU) to represent continuous memory trajectories via orthogonal basis functions. This architecture significantly reduces computational complexity while enhancing responsiveness to non‐linear memory changes. Furthermore, a PPO‐MMC algorithm is developed to optimize review intervals within a continuous state space. By achieving joint learning of memory prediction and policy scheduling, the framework effectively minimizes review costs while maximizing long‐term retention. This paper validated through ablation and comparative experiments that the mean absolute error (MAE) of the GLD‐HLR model's recall probability predictions remained below 0.03, achieving at least a 4% reduction compared to the LSTM‐HLR model. The mean absolute percentage error (MAPE) for half‐life predictions was below 0.2, which is smaller than the prediction errors of all other models. The PPO‐MMC algorithm achieved a cumulative number of words learned (WTL) exceeding 8000 within 1000 days, with the number of words memorized at the target half‐life (THR) surpassing 7000. This indicates that the algorithm can efficiently help learners master a large number of vocabulary words within a limited time frame and achieve long‐term retention.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhao et al. (2026) studied this question.

synapsesocial.com/papers/6a211852d499ed480b170eb8https://doi.org/10.1111/exsy.70269
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Long Short-Term Memory1997 · 101,723 citations
  2. 2DRL-SRS: A Deep Reinforcement Learning Approach for Optimizing Spaced Repetition Scheduling2024 · 6 citations
  3. 3Deep Reinforcement Learning (DRL): Another Perspective for Unsupervised Wireless Localization2019 · 146 citations
  4. 4A Review of Recurrent Neural Networks: LSTM Cells and Network Architectures2019 · 5,649 citations
  5. 5The Impact of Spaced Repetition Learning on the Learning Success in Mobile Learning Games2021 · 9 citations