What does this research mean for the field?

A multiple model-based reinforcement learning (MMRL) architecture that decomposes complex tasks based on environmental predictability effectively solves nonlinear, nonstationary control tasks in both discrete and continuous environments. Novelty: ClaimNovelty.METHODOLOGICAL. Consensus alignment: ConsensusAlignment.NEUTRAL.

June 1, 2002

Multiple Model-Based Reinforcement Learning

Key Points

Key points are not available for this paper at this time.

Abstract

We propose a modular reinforcement learning architecture for nonlinear, nonstationary control tasks, which we call multiple model-based reinforcement learning (MMRL). The basic idea is to decompose a complex task into multiple domains in space and time based on the predictability of the environmental dynamics. The system is composed of multiple modules, each of which consists of a state prediction model and a reinforcement learning controller. The "responsibility signal," which is given by the softmax function of the prediction errors, is used to weight the outputs of multiple modules, as well as to gate the learning of the prediction models and the reinforcement learning controllers. We formulate MMRL for both discrete-time, finite-state case and continuous-time, continuous-state case. The performance of MMRL was demonstrated for discrete case in a nonstationary hunting task in a grid world and for continuous case in a nonlinear, nonstationary control task of swinging up a pendulum with variable physical parameters.

Connected Papers

Building similarity graph...

Analyzing shared references across papers

Discussion

Authors

Kenji Doya

Okinawa Institute of Science and Technology Graduate University

Kazuyuki Samejima

Tamagawa University

Ken-ichi Katagiri

Nara Institute of Science and Technology

Journals

Neural Computation

Actions

Institutions

Nara Institute of Science and Technology

Kyoto Seika University

References and Citations

Connected Papers

Building similarity graph...

Analyzing shared references across papers

Discussion

Cite this study

Doya et al. (Sat,) studied this question.

synapsesocial.com/papers/6a07a125d343c0cd6cc63e63 — DOI: https://doi.org/10.1162/089976602753712972

Also consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

Adaptive Mixtures of Local Experts· 1991 · 4,951 citations
10. Reinforcement learning· 2020 · 1,950 citations
Reinforcement Learning· 1998 · 3,005 citations
Introduction to Reinforcement Learning· 1998 · 6,922 citations
MOSAIC Model for Sensorimotor Learning and Control· 2001 · 730 citations

Multiple Model-Based Reinforcement Learning

Key Points

Abstract

Citation Network

Connected Papers

Discussion

Authors

Journals

Actions

Institutions

References and Citations

Citation Network

Connected Papers

Discussion

Cite this study

Also consider

Also consider