Quasi-Stochastic Approximation and Off-Policy Reinforcement Learning

Quasi-Stochastic Approximation and Off-Policy Reinforcement Learning
复制标题

准随机逼近和离策略强化学习

DOI:
10.1109/cdc40024.2019.9029247
复制
发表时间:
2019
期刊:
Proceedings of the IEEE Conference on Decision Control
影响因子:
--
通讯作者:
Meyn, Sean
Meyn, Sean
中科院分区:
--
文献类型:
--
作者:
Bernstein, Andrey;Chen, Yue;Colombino, Marcello;Dall'Anese, Emiliano;Mehta, Prashant;Meyn, Sean

文献摘要

参考文献

被引文献

相似文献

多元罗宾斯-门罗过程的牛顿-拉夫森版本
DOI: 10.1214/aos/1176346589
发表时间: 1985
影响因子: 4.5
作者:
D. Ruppert
通讯作者: D. Ruppert
DOI: --
发表时间: 2019
期刊:
影响因子: --
作者:
A. Bernstein;Yue;Marcello Colombino;E. Dall’Anese;P. Mehta;Sean P. Meyn
通讯作者: Sean P. Meyn
DOI: --
发表时间: 2011
期刊: Proceedings of the 2011 American Control Conference
影响因子: --
作者:
Darshan Shirodkar;Sean P. Meyn
通讯作者: Sean P. Meyn
DOI: 10.1109/jiot.2018.2839563
发表时间: 2017-07
影响因子: 10.6
作者:
Tianyi Chen;G. Giannakis
通讯作者: Tianyi Chen;G. Giannakis
DOI: --
发表时间: 2012
期刊:
影响因子: --
作者:
Shu;M. Krstić
通讯作者: M. Krstić