错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Walking Motion Generation of Bipedal Robot Based on Planar Covariation Using Deep Reinforcement Learning

  • Junsei Yamano,
  • Masaki Kurokawa,
  • Yuki Sakai,
  • Kenji Hashimoto

摘要

Recently, a deep reinforcement learning approach has been proposed as one of the control methods for legged robots. In this paper, we describe the walking motion of a bipedal robot using deep reinforcement learning and define a human-like biped gait as a gait in which the robot lands on its heels and takes off its toes. In this study, we focus on the kinematic synergy (planar covariation) pointed out by Borghese et al. as a characteristic gait of humans. Planar covariation is angular data plotted in the sagittal plane with the elevation angles at the thigh, shank, and foot on the three axes falling on a single plane. We propose this feature as the reward for reinforcement learning. By introducing this reward, the bipedal robot realized a human-like gait. We also consider that the proposed reward is one factor that characterizes human walking motion by comparing the learning results when this feature is not used.