A simple rule how to make a reward for learning with human interaction
A simple rule how to make a reward for learning with human interaction
复制标题
一个简单的规则,如何通过人际互动来奖励学习
DOI:
10.1109/cira.2007.382921
复制
发表时间:
2007
期刊:
影响因子:
--
通讯作者:
K. Kurashige
中科院分区:
文献类型:
--
作者:
K. Kurashige
Various learning methods are adapted for experimental robot. We can make movement of a robot by giving teaching signals to a robot. But it is heavy for operator to define how to give teaching signals generally because operator must guess and think of a task and environment and define a function to do that. Here the author aim to create teaching signals automatically for each task and environment. In this paper, the author suggest a simple rule which is independent of information about any task and environment to create teaching signals for each task and environment. This rule is that a situation which is often happened is good situation. In this paper, the author adopt reinforcement learning as learning method and a small-sized humanoid robot as application. The author show creating a reward by adapting a rule and show that a robot can learn and make movement.