PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 30, 20240 citationsOpen Access

Q-learning as a monotone scheme

View Full Paper
LYLingyi Yang

Key Points

Key points are not available for this paper at this time.

Abstract

Stability issues with reinforcement learning methods persist. To better understand some of these stability and convergence issues involving deep reinforcement learning methods, we examine a simple linear quadratic example. We interpret the convergence criterion of exact Q-learning in the sense of a monotone scheme and discuss consequences of function approximation on monotonicity properties.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Lingyi Yang (2024) studied this question.

synapsesocial.com/papers/68e67cafb6db64358760643fhttps://doi.org/10.48550/arxiv.2405.20538
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms2024 · 1 citations
  2. 2TD-Learning and Q-Learning: A Survey of Theory, Analysis, and Trends2026
  3. 3Two-Step Q-Learning2024
  4. 4When Q-Learning fails: unstable behavior for infinite state spaces2026
  5. 5On the Stability of Learning in Network Games with Many Players2024