Mean-Variance Tradeoffs in an Undiscounted MDP: The Unichain Case

Mean-Variance Tradeoffs in an Undiscounted MDP: The Unichain Case
复制标题

DOI:
10.1287/opre.42.1.184
复制
发表时间:
1994-02
期刊:
Oper. Res.
影响因子:
--
通讯作者:
Kun-Jen Chung
Kun-Jen Chung
中科院分区:
其他
文献类型:
--
作者:
Kun-Jen Chung

文献摘要

被引文献

相似文献

这里分析的问题是计算Pareto最优的意义下的高均值和低方差的平稳分布的单链,非折扣马尔可夫决策过程MDP,简称。
The problem analyzed here is the computation of Pareto optima in the sense of high mean and low variance of the stationary distribution in the unichain, undiscounted Markov decision process MDP, for short.