Imitation Learning With Stability and Safety Guarantees
Imitation Learning With Stability and Safety Guarantees
复制标题
DOI:
10.1109/lcsys.2021.3077861
复制
发表时间:
2020-12
影响因子:
3
通讯作者:
He Yin;P. Seiler;Ming Jin;M. Arcak
中科院分区:
文献类型:
--
作者:
He Yin;P. Seiler;Ming Jin;M. Arcak
A method is presented to learn neural network (NN) controllers with stability and safety guarantees through imitation learning (IL). Convex stability and safety conditions are derived for linear time-invariant systems with NN controllers by merging Lyapunov theory with local quadratic constraints to bound the activation functions in the NN. These conditions are incorporated in the IL process, which minimizes the IL loss, and maximizes the volume of the region of attraction associated with the NN controller simultaneously. An alternating direction method of multipliers based algorithm is proposed to solve the IL problem. The method is illustrated on a vehicle lateral control example.