Speeding-up Reinforcement Learning with Multi-step Actions
Speeding-up Reinforcement Learning with Multi-step Actions
复制标题
通过多步骤行动加速强化学习
DOI:
10.1007/3-540-46084-5_132
复制
发表时间:
2002
期刊:
影响因子:
--
通讯作者:
Martin A. Riedmiller
中科院分区:
文献类型:
--
作者:
Ralf Schoknecht;Martin A. Riedmiller
In recent years hierarchical concepts of temporal abstraction have been integrated in the reinforcement learning framework to improve scalability. However, existing approaches are limited to domains where a decomposition into subtasks is known a priori. In this paper we propose the concept of explicitly selecting time scale related actions if no subgoalrelated abstract actions are available. This is realised with multistep actions on different time scales that are combined in one single action set. The special structure of the action set is exploited in the MSAQ-learning algorithm. By learning on different explicitly specified time scales simultaneously, a considerable improvement of learning speed can be achieved. This is demonstrated on two benchmark problems.