CPG-RL: Learning Central Pattern Generators for Quadruped Locomotion

In this letter, we present a method for integrating central pattern generators (CPGs), i.e. systems of coupled oscillators, into the deep reinforcement learning (DRL) framework to produce robust and omnidirectional quadruped locomotion. The agent learns to directly modulate the intrinsic oscillator setpoints (amplitude and frequency) and coordinate rhythmic behavior among different oscillators. This approach also allows the use of DRL to explore questions related to neuroscience, namely the role of descending pathways, interoscillator couplings, and sensory feedback in gait generation. We train our policies in simulation and perform a sim-to-real transfer to the Unitree A1 quadruped, where we observe robust behavior to disturbances unseen during training, most notably to a dynamically added 13.75 kg load representing 115% of the nominal quadruped mass. We test several different observation spaces based on proprioceptive sensing and show that our framework is deployable with no domain randomization and very little feedback, where along with the oscillator states, it is possible to provide only contact booleans in the observation space.

Chattez avec Graph Search

Posez n’importe quelle question sur les cours, conférences, exercices, recherches, actualités, etc. de l’EPFL ou essayez les exemples de questions ci-dessous.

AVERTISSEMENT : Le chatbot Graph n'est pas programmé pour fournir des réponses explicites ou catégoriques à vos questions. Il transforme plutôt vos questions en demandes API qui sont distribuées aux différents services informatiques officiellement administrés par l'EPFL. Son but est uniquement de collecter et de recommander des références pertinentes à des contenus que vous pouvez explorer pour vous aider à répondre à vos questions.

CPG-RL: Learning Central Pattern Generators for Quadruped Locomotion

Graph Chatbot

Chattez avec Graph Search

Fusing Pre-existing Knowledge and Machine Learning for Enhanced Building Thermal Modeling and Control

Inverse design of metal-organic frameworks for direct air capture of CO2via deep reinforcement learning

Multi-agent Reinforcement Learning for Assembly of a Spanning Structure

Multi-agent Reinforcement Learning for Assembly of a Spanning Structure

Inverse design of metal-organic frameworks for direct air capture of CO2via deep reinforcement learning

Fusing Pre-existing Knowledge and Machine Learning for Enhanced Building Thermal Modeling and Control