PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
June 20, 2007185 citationsOpen Access

Reinforcement learning by reward-weighted regression for operational space control

JPJan PetersSSStefan Schaal

Key Points

Key points are not available for this paper at this time.

Abstract

Many robot control problems of practical importance, including operational space control, can be reformulated as immediate reward reinforcement learning problems. However, few of the known optimization or reinforcement learning algorithms can be used in online learning control for robots, as they are either prohibitively slow, do not scale to interesting domains of complex robots, or require trying out policies generated by random search, which are infeasible for a physical system. Using a generalization of the EM-base reinforcement learning framework suggested by Dayan Hinton, we reduce the problem of learning with immediate rewards to a reward-weighted regression problem with an adaptive, integrated reward transformation for faster convergence. The resulting algorithm is efficient, learns smoothly without dangerous jumps in solution space, and works well in applications of complex high degreeof-freedom robots. 1.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Peters et al. (2007) studied this question.

synapsesocial.com/papers/6a0f6d5528c2d29469fe193dhttps://doi.org/10.1145/1273496.1273590
Ask AI
Helpful
Bookmark
Share
View Full Paper