Empirical Algorithms for General Stochastic Systems with Continuous States and Actions
Empirical Algorithms for General Stochastic Systems with Continuous States and Actions
复制标题
DOI:
10.1109/cdc40024.2019.9029308
复制
发表时间:
2019-12
期刊:
影响因子:
--
通讯作者:
Hiteshi Sharma;R. Jain;W. Haskell
中科院分区:
文献类型:
--
作者:
Hiteshi Sharma;R. Jain;W. Haskell
In this paper, we present Randomized Empirical Value Learning (RAEVL) algorithm for MDPs with continuous state and action spaces. This algorithm combines the ideas of random search over action space with randomized function approximation method to generalize the value functions over state space . Our theoretical analysis is done under a random operator framework combined with stochastic dominance argument. This provides finite-time analysis of the proposed algorithm as well as give the sample complexity.