PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
June 2, 2021465 citationsOpen Access

Decision Transformer: Reinforcement Learning via Sequence Modeling

View Full Paper
LCLili ChenKLKevin LüARAravind Rajeswaran

Key Points

Key points are not available for this paper at this time.

Abstract

We introduce a framework that abstracts Reinforcement Learning (RL) as a sequence modeling problem. This allows us to draw upon the simplicity and scalability of the Transformer architecture, and associated advances in language modeling such as GPT-x and BERT. In particular, we present Decision Transformer, an architecture that casts the problem of RL as conditional sequence modeling. Unlike prior approaches to RL that fit value functions or compute policy gradients, Decision Transformer simply outputs the optimal actions by leveraging a causally masked Transformer. By conditioning an autoregressive model on the desired return (reward), past states, and actions, our Decision Transformer model can generate future actions that achieve the desired return. Despite its simplicity, Decision Transformer matches or exceeds the performance of state-of-the-art model-free offline RL baselines on Atari, OpenAI Gym, and Key-to-Door tasks.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Chen et al. (2021) studied this question.

synapsesocial.com/papers/69d8f85bade63f05b9bee015https://doi.org/10.48550/arxiv.2106.01345
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Hafez: an Interactive Poetry Generation System2017 · 135 citations
  2. 2GeDi: Generative Discriminator Guided Sequence Generation2021 · 215 citations
  3. 3Reset-Free Lifelong Learning with Skill-Space Planning2020 · 11 citations