Introduction of Fixed Mode States into Online Reinforcement Learning with Penalty and Reward and Its Application to Waist Trajectory Generation of Biped Robot

Introduction of Fixed Mode States into Online Reinforcement Learning with Penalty and Reward and Its Application to Waist Trajectory Generation of Biped Robot
复制标题

将固定模式状态引入带惩罚和奖励的在线强化学习及其在双足机器人腰部轨迹生成中的应用

DOI:
--
复制
发表时间:
2012
影响因子:
0.7
通讯作者:
Kazuteru Miyazaki and Hiroaki Kobayashi
Kazuteru Miyazaki and Hiroaki Kobayashi
中科院分区:
--
文献类型:
--
作者:
Seiya Kuroda;Kazuteru Miyazaki and Hiroaki Kobayashi

文献摘要

相似文献