An off-policy natural gradient method for a partial observable Markov decision process
An off-policy natural gradient method for a partial observable Markov decision process
复制标题
部分可观测马尔可夫决策过程的离策略自然梯度法
DOI:
--
复制
发表时间:
2005
期刊:
影响因子:
--
通讯作者:
Y.
中科院分区:
文献类型:
--
作者:
Nakamura;Y.
DOI:
--
发表时间:
2003
期刊:
Artificial Neural Networks and Neural Information Processing, Lecture Notes in Computer Science (Berlin : Springer-Verlag) 2714
影响因子:
--
作者:
Yoshimoto;J.
通讯作者:
J.
DOI:
--
发表时间:
2004
期刊:
Parallel Problem Solving from Nature - PPSN VIII, Lecture Notes in Computer Science 3242
影响因子:
--
作者:
Nakamura;Y.
通讯作者:
Y.