Implementation of deep deterministic policy gradients for controlling dynamic bipedalwalking

13Citations
Citations of this article
31Readers
Mendeley users who have this article in their library.

Abstract

A control system for bipedal walking in the sagittal plane was developed in simulation. The biped model was built based on anthropometric data for a 1.8 m tall male of average build. At the core of the controller is a deep deterministic policy gradient (DDPG) neural network that was trained in GAZEBO, a physics simulator, to predict the ideal foot placement to maintain stable walking despite external disturbances. The complexity of the DDPG network was decreased through carefully selected state variables and a distributed control system. Additional controllers for the hip joints during their stance phases and the ankle joint during toe-off phase help to stabilize the biped during walking. The simulated biped can walk at a steady pace of approximately 1 m/s, and during locomotion it can maintain stability with a 30 kgm/s impulse applied forward on the torso or a 40 kgm/s impulse applied rearward. It also maintains stable walking with a 10 kg backpack or a 25 kg front pack. The controller was trained on a 1.8 m tall model, but also stabilizes models 1.4-2.3 m tall with no changes.

Cite

CITATION STYLE

APA

Liu, C., Lonsberry, A. G., Nandor, M. J., Audu, M. L., Lonsberry, A. J., & Quinn, R. D. (2019). Implementation of deep deterministic policy gradients for controlling dynamic bipedalwalking. Biomimetics, 4(1). https://doi.org/10.3390/biomimetics4010028

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free