Introduction of Fixed Mode States into Online Reinforcement Learning with Penalty and Reward and Its Application to Waist Trajectory Generation of Biped Robot
Introduction of Fixed Mode States into Online Reinforcement Learning with Penalty and Reward and Its Application to Waist Trajectory Generation of Biped Robot
复制标题
将固定模式状态引入带惩罚和奖励的在线强化学习及其在双足机器人腰部轨迹生成中的应用
DOI:
--
复制
发表时间:
2012
影响因子:
0.7
通讯作者:
Kazuteru Miyazaki and Hiroaki Kobayashi
中科院分区:
文献类型:
--
作者:
Seiya Kuroda;Kazuteru Miyazaki and Hiroaki Kobayashi