PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 21, 20241 citationsOpen Access

Bayesian Optimization for Sample-Efficient Policy Improvement in Robotic Manipulation

View Full Paper
ARAdrian RöferINIman NematollahiTWTim Welschehold

Key Points

  • Sample efficiency improves with the BOpt-GMM approach, which combines imitation and experience learning.
  • Demonstrated effectiveness through simulations and real-world experiments, with a focus on complex tasks.
  • Assessment using a hybrid model that utilizes a Gaussian Mixture Model to refine skills over time and interactions leads to improved performance across trials and environments effectively.. The Bayesian optimization builds on limited data, addressing learning challenges in robotics efficiently.

Abstract

Sample efficient learning of manipulation skills poses a major challenge in robotics. While recent approaches demonstrate impressive advances in the type of task that can be addressed and the sensing modalities that can be incorporated, they still require large amounts of training data. Especially with regard to learning actions on robots in the real world, this poses a major problem due to the high costs associated with both demonstrations and real-world robot interactions. To address this challenge, we introduce BOpt-GMM, a hybrid approach that combines imitation learning with own experience collection. We first learn a skill model as a dynamical system encoded in a Gaussian Mixture Model from a few demonstrations. We then improve this model with Bayesian optimization building on a small number of autonomous skill executions in a sparse reward setting. We demonstrate the sample efficiency of our approach on multiple complex manipulation skills in both simulations and real-world experiments. Furthermore, we make the code and pre-trained models publicly available at http://bopt-gmm. cs.uni-freiburg.de.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Röfer et al. (2024) studied this question.

synapsesocial.com/papers/68e7309eb6db6435876aa7c8https://doi.org/10.48550/arxiv.2403.14305
Ask AI
Helpful
Bookmark
Share
View Full Paper