Bandits all the way down: UCB1 as a simulation policy in Monte Carlo Tree Search

Bandits all the way down: UCB1 as a simulation policy in Monte Carlo Tree Search
复制标题

一路强盗:UCB1 作为蒙特卡罗树搜索中的模拟策略

DOI:
10.1109/cig.2013.6633613
复制
发表时间:
2013
期刊:
--
影响因子:
--
通讯作者:
Powley E
Powley E
中科院分区:
--
文献类型:
--
作者:
Powley E

文献摘要

参考文献

被引文献

相似文献

多人蒙特卡罗树搜索的增强功能
DOI: --
发表时间: 2010
期刊: Computers and Games
影响因子: --
作者:
J. A. M. Nijssen;M. Winands;H. J. Herik;H. Iida;A. Plaat
通讯作者: A. Plaat
Ieee Transactions on Computational Intelligence and Ai in Games 1 N-grams 和一般游戏中应用的最后良好回复策略
DOI: --
发表时间: --
期刊:
影响因子: --
作者:
Mandy J. W. Tak;M. Winands;Y. Björnsson
通讯作者: Y. Björnsson
哈瓦那的蒙特卡洛树搜索增强功能
DOI: --
发表时间: 2011
期刊: Advances in Computer Games
影响因子: --
作者:
J. Stankiewicz;M. Winands;J. Uiterwijk
通讯作者: J. Uiterwijk
遗忘的力量:改进 Monte Carlo Go 中的最后良好回复策略
DOI: --
发表时间: 2010
影响因子: --
作者:
Hendrik Baier;P. Drake
通讯作者: P. Drake
同步比赛中 UCT 与 CFR 的比较
DOI: --
发表时间: 2009
期刊:
影响因子: --
作者:
Mohammad Shafiei;Nathan R Sturtevant;J. Schaeffer
通讯作者: J. Schaeffer