Reinforcement learning for adaptive optimal control of continuous-time linear periodic systems
Reinforcement learning for adaptive optimal control of continuous-time linear periodic systems
复制标题
DOI:
10.1016/j.automatica.2020.109035
复制
发表时间:
2020-08
期刊:
影响因子:
--
通讯作者:
Bo Pang;Zhong-Ping Jiang;I. Mareels
中科院分区:
文献类型:
--
作者:
Bo Pang;Zhong-Ping Jiang;I. Mareels
This paper studies the infinite-horizon adaptive optimal control of continuous-time linear periodic (CTLP) systems, using reinforcement learning techniques. By means of policy iteration (PI) for CTLP systems, both on-policy and off-policy adaptive dynamic programming (ADP) algorithms are derived, such that the solution of the optimal control problem can be found without the exact knowledge of the system dynamics. Starting with initial stabilizing controllers, the proposed PI-based ADP algorithms converge to the optimal solutions under mild conditions. Application to the adaptive optimal control of the lossy Mathieu equation demonstrates the efficacy of the proposed learning-based adaptive optimal control algorithm.