Physics-Model-Regulated Deep Reinforcement Learning Towards Safety & Stability Guarantees
Physics-Model-Regulated Deep Reinforcement Learning Towards Safety & Stability Guarantees
复制标题
DOI:
10.1109/cdc49753.2023.10383560
复制
发表时间:
2023-12
期刊:
影响因子:
--
通讯作者:
Hongpeng Cao;Yanbing Mao;Lui Sha;Marco Caccamo
中科院分区:
文献类型:
--
作者:
Hongpeng Cao;Yanbing Mao;Lui Sha;Marco Caccamo
Deep reinforcement learning (DRL) has demonstrated impressive success in solving complex control tasks by synthesizing control policies from data. However, the safety and stability of applying DRL to safety-critical systems remain a primary concern and challenging problem. To address the problem, we propose the Phy-DRL: a novel physics-model-regulated deep reinforcement learning framework. The Phy-DRL is novel in two architectural designs: a physics-model-regulated reward and residual control, which integrates physics-model-based control and data-driven control. The concurrent designs enable the Phy-DRL the mathematically provable safety and stability guarantees. Finally, the effectiveness of the Phy-DRL is validated by an inverted pendulum system. Additionally, the experimental results demonstrate that the Phy-DRL features remarkably accelerated training and enlarged reward.