PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 29, 20250 citationsOpen Access

Gait in Eight: Efficient On-Robot Learning for Omnidirectional Quadruped Locomotion

View Full Paper
NBNico BohlingerJKJonathan KinzelDPDaniel Palenicek

Key Points

  • Our framework demonstrates effective locomotion learning in just 8 minutes of real-time training.
  • Utilizing the off-policy algorithm CrossQ, we achieve high sample efficiency in quadruped training.
  • The approach combines predictive models and control architectures to enhance both speed and stability.
  • Results validate the method across varied environments, indicating robustness and adaptability.

Abstract

On-robot Reinforcement Learning is a promising approach to train embodiment-aware policies for legged robots. However, the computational constraints of real-time learning on robots pose a significant challenge. We present a framework for efficiently learning quadruped locomotion in just 8 minutes of raw real-time training utilizing the sample efficiency and minimal computational overhead of the new off-policy algorithm CrossQ. We investigate two control architectures: Predicting joint target positions for agile, high-speed locomotion and Central Pattern Generators for stable, natural gaits. While prior work focused on learning simple forward gaits, our framework extends on-robot learning to omnidirectional locomotion. We demonstrate the robustness of our approach in different indoor and outdoor environments.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Bohlinger et al. (2025) studied this question.

synapsesocial.com/papers/68da58c9c1728099cfd10a39https://doi.org/10.48550/arxiv.2503.08375
Ask AI
Helpful
Bookmark
Share
View Full Paper