Experimental Demonstration of Adaptive MDP-Based Planning with Model Uncertainty
Experimental Demonstration of Adaptive MDP-Based Planning with Model Uncertainty
复制标题
具有模型不确定性的基于自适应 MDP 规划的实验演示
DOI:
--
复制
发表时间:
2008
期刊:
影响因子:
--
通讯作者:
J. How
中科院分区:
文献类型:
--
作者:
Brett Bethke;L. Bertuccelli;J. How
Markov decision processes (MDPs) are a natural framework for solving multiagent planning problems since they can model stochastic system dynamics and interdependencies between agents. In these approaches, accurate modeling of the system in question is important, since mismodeling may lead to severely degraded performance (i.e. loss of vehicles). Furthermore, in many problems of interest, it may be dicult or impossible to obtain an accurate model before the system begins operating; rather, the model must be estimated online. Therefore, an adaptation mechanism that can estimate the system model and adjust the system control policy online can improve performance over a static (o-line) approach. This paper presents an MDP formulation of a multi-agent persistent surveillance problem and shows, in simulation, the importance of accurate modeling of the system. An adaptation mechanism, consisting of a Bayesian model estimator and a continuouslyrunning MDP solver, is then discussed. Finally, we present hardware flight results from the MIT RAVEN testbed that clearly demonstrate the performance benefits of this adaptive approach in the persistent surveillance problem.