首页 /研究 /Learning Sensory Feedback to CPG with Policy Gradient for Biped Locomotion
LOCOMOTION

Learning Sensory Feedback to CPG with Policy Gradient for Biped Locomotion

Takamitsu Matsubara, Jun Morimoto, Jun Nakanishi, Masa-aki Sato, Kenji Doya

发表年份
2006
引用次数
30

摘要

This paper proposes a learning framework for a CPG-based biped locomotion controller using a policy gradient method. Our goal in this study is to develop an efficient learning algorithm by reducing the dimensionality of the state space used for learning. We demonstrate that an appropriate feedback controller in the CPG-based controller can be acquired using the proposed method within a few thousand trials by numerical simulations. Furthermore, we implement the learned controller on the physical biped robot to experimentally show that the learned controller successfully works in the real environment.

关键词

Controller (irrigation)Computer scienceControl theory (sociology)RobotCurse of dimensionalityReinforcement learningControl engineeringArtificial intelligenceEngineeringControl (management)

相关论文

查看 LOCOMOTION 分类全部论文