PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 2006188 citations

Autonomous shaping

View Full Paper
GKGeorge KonidarisABAndrew G. Barto

Key Points

Key points are not available for this paper at this time.

Abstract

We introduce the use of learned shaping rewards in reinforcement learning tasks, where an agent uses prior experience on a sequence of tasks to learn a portable predictor that estimates intermediate rewards, resulting in accelerated learning in later tasks that are related but distinct. Such agents can be trained on a sequence of relatively easy tasks in order to develop a more informative measure of reward that can be transferred to improve performance on more difficult tasks without requiring a hand coded shaping function. We use a rod positioning task to show that this significantly improves performance even after a very brief training period.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Konidaris et al. (2006) studied this question.

synapsesocial.com/papers/6a1d557633e2df9c962f579chttps://doi.org/10.1145/1143844.1143906
Ask AI
Helpful
Bookmark
Share
View Full Paper