PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 19, 20240 citationsOpen Access

Yell At Your Robot: Improving On-the-Fly from Language Corrections

View Full Paper
LSLucy Xiaoyang ShiZHZheyuan HuTZTony Z. Zhao

Key Points

Key points are not available for this paper at this time.

Abstract

Hierarchical policies that combine language and low-level control have been shown to perform impressively long-horizon robotic tasks, by leveraging either zero-shot high-level planners like pretrained language and vision-language models (LLMs/VLMs) or models trained on annotated robotic demonstrations. However, for complex and dexterous skills, attaining high success rates on long-horizon tasks still represents a major challenge -- the longer the task is, the more likely it is that some stage will fail. Can humans help the robot to continuously improve its long-horizon task performance through intuitive and natural feedback? In this paper, we make the following observation: high-level policies that index into sufficiently rich and expressive low-level language-conditioned skills can be readily supervised with human feedback in the form of language corrections. We show that even fine-grained corrections, such as small movements ("move a bit to the left"), can be effectively incorporated into high-level policies, and that such corrections can be readily obtained from humans observing the robot and making occasional suggestions. This framework enables robots not only to rapidly adapt to real-time language feedback, but also incorporate this feedback into an iterative training scheme that improves the high-level policy's ability to correct errors in both low-level execution and high-level decision-making purely from verbal feedback. Our evaluation on real hardware shows that this leads to significant performance improvement in long-horizon, dexterous manipulation tasks without the need for any additional teleoperation. Videos and code are available at https://yay-robot.github.io/.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Shi et al. (2024) studied this question.

synapsesocial.com/papers/68e7362fb6db6435876b0352https://doi.org/10.48550/arxiv.2403.12910
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Learning to Learn Faster from Human Feedback with Language Model Predictive Control2024 · 1 citations
  2. 2Enhancing stability and reliability in LLM-driven robotic manipulation through human skill demonstration and visual tracking2026 · 1 citations
  3. 3LGR2: Language Guided Reward Relabeling for Accelerating Hierarchical Reinforcement Learning2024
  4. 4Self-Corrected Multimodal Large Language Model for End-to-End Robot Manipulation2024 · 2 citations
  5. 5Neuro-symbolic Hierarchical Learning for Long-Horizon Robotic Tasks2026