Approximate Relative Value Learning for Average-reward Continuous State MDPs
Approximate Relative Value Learning for Average-reward Continuous State MDPs
复制标题
平均奖励连续状态 MDP 的近似相对价值学习
DOI:
--
复制
发表时间:
2019
期刊:
影响因子:
--
通讯作者:
Jain, Rahul
中科院分区:
文献类型:
--
作者:
Sharma, Hiteshi;Jafarnia-Jahromi, Mehdi;Jain, Rahul
DOI:
10.1109/tit.1978.1055865
发表时间:
1978
期刊:
IEEE Trans. Inf. Theory
影响因子:
--
作者:
L. Devroye
通讯作者:
L. Devroye
DOI:
10.1007/978-3-540-75225-7_30
发表时间:
2007
期刊:
2019 18th European Control Conference (ECC)
影响因子:
--
作者:
R. Ortner
通讯作者:
R. Ortner
DOI:
10.23919/ecc.2019.8795982
发表时间:
2019
期刊:
2019 18th European Control Conference (ECC
影响因子:
--
作者:
Sharma, Hiteshi;Jain, Rahul;Gupta, Abhishek
通讯作者:
Gupta, Abhishek