PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 27, 20241 citationsOpen Access

Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning

View Full Paper
XZXiaoyu ZhangMCMatthew S. ChangPKPranav Kumar

Key Points

Key points are not available for this paper at this time.

Abstract

A common failure mode for policies trained with imitation is compounding execution errors at test time. When the learned policy encounters states that were not present in the expert demonstrations, the policy fails, leading to degenerate behavior. The Dataset Aggregation, or DAgger approach to this problem simply collects more data to cover these failure states. However, in practice, this is often prohibitively expensive. In this work, we propose Diffusion Meets DAgger (DMD), a method to reap the benefits of DAgger without the cost for eye-in-hand imitation learning problems. Instead of collecting new samples to cover out-of-distribution states, DMD uses recent advances in diffusion models to create these samples with diffusion models. This leads to robust performance from few demonstrations. In experiments conducted for non-prehensile pushing on a Franka Research 3, we show that DMD can achieve a success rate of 80% with as few as 8 expert demonstrations, where naive behavior cloning reaches only 20%. DMD also outperform competing NeRF-based augmentation schemes by 50%.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhang et al. (2024) studied this question.

synapsesocial.com/papers/68e77692b6db6435876eb575https://doi.org/10.48550/arxiv.2402.17768
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Uncertainty-Driven Data Aggregation for Imitation Learning in Autonomous Vehicles2024
  2. 2Behavioral Refinement via Interpolant-based Policy Diffusion2024 · 1 citations
  3. 3Diffusion Augmented Agents: A Framework for Efficient Exploration and Transfer Learning2024 · 1 citations
  4. 4TubeDAgger: Reducing the Number of Expert Interventions with Stochastic Reach-Tubes2025
  5. 5DiffAIL: Diffusion Adversarial Imitation Learning2024 · 9 citations