Instance-Based Policy Search using Binomial Distribution Crossover and Iterated Refreshment
Instance-Based Policy Search using Binomial Distribution Crossover and Iterated Refreshment
复制标题
使用二项分布交叉和迭代刷新的基于实例的策略搜索
DOI:
10.1109/cec.2006.1688333
复制
发表时间:
2006
期刊:
影响因子:
--
通讯作者:
S. Kobayashi
中科院分区:
文献类型:
--
作者:
Chikao Tsuchiya;Kokolo Ikeda;J. Sakuma;I. Ono;S. Kobayashi
This paper describes a GA based lazy approach toward reinforcement learning. This approach employs data-driven policy, which is composed of an instance set and an instance-based action selector. This feature provides a number of advantages. However some difficulties remain uninvestigated. One of them is the huge and complicated search space. We have an idea that preserving characteristics of the GA population and introducing new characteristics can overcome these difficulties. On the basis of this idea, we propose two genetic operators; Binomial Distribution Crossover (BDX) and iterated refreshment. The BDX generates the descendants inheriting the parents' characteristics and the iterated refreshment introduces new characteristics greedily. The GA powered by these operators was applied to the benchmark tasks to demonstrate the ability. Each operator also was investigated and discussed from the various perspectives. Finally, we provide the preferable parameter settings for our method.
影响因子:
1.6
作者:
J. Santamaría;R. Sutton;A. Ram
通讯作者:
J. Santamaría;R. Sutton;A. Ram