Coordination of multiple behaviors acquired by a vision-based reinforcement learning
Coordination of multiple behaviors acquired by a vision-based reinforcement learning
复制标题
基于视觉的强化学习获得的多种行为的协调
DOI:
10.1109/iros.1994.407484
复制
发表时间:
1994
期刊:
影响因子:
--
通讯作者:
K. Hosoda
中科院分区:
文献类型:
--
作者:
M. Asada;E. Uchibe;S. Noda;Sukoya Tawaratsumida;K. Hosoda
A method is proposed which accomplishes a whole task consisting of plural subtasks by coordinating multiple behaviors acquired by a vision-based reinforcement learning. First, individual behaviors which achieve the corresponding subtasks are independently acquired by Q-learning, a widely used reinforcement learning method. Each learned behavior can be represented by an action-value function in terms of state of the environment and robot action. Next, three kinds of coordinations of multiple behaviors are considered; simple summation of different action-value functions, switching action-value functions according to situations, and learning with previously obtained action-value functions as initial values of a new action-value function. A task of shooting a ball into the goal avoiding collisions with an enemy is examined. The task can be decomposed into a ball shooting subtask and a collision avoiding subtask. These subtasks should be accomplished simultaneously, but they are not independent of each other.<<ETX>>