Optimal Control-Based Adaptive NN Design for a Class of Nonlinear Discrete-Time Block-Triangular Systems

Optimal Control-Based Adaptive NN Design for a Class of Nonlinear Discrete-Time Block-Triangular Systems
复制标题

一类非线性离散时间块三角系统的基于最优控制的自适应神经网络设计

DOI:
10.1109/tcyb.2015.2494007
复制
发表时间:
2016-02
影响因子:
11.8
通讯作者:
Tong Shaocheng
Tong Shaocheng
中科院分区:
计算机科学1区
文献类型:
--
作者:
Liu Yan-Jun;Tong Shaocheng

文献摘要

参考文献

被引文献

相似文献

针对一类未知非线性离散时间系统,提出了一种基于最优控制方案的自适应神经网络设计方法。被控系统是块三角形多输入多输出纯反馈结构,即,在每个子系统的每个方程中都包括状态和输入耦合以及非仿射函数。设计的目标是提供一种控制方案,既保证系统的稳定性,又达到最优的控制性能。本文的主要贡献是首次实现了这类系统的最优性能。由于子系统之间的相互作用,使最优控制信号是一个困难的任务。设计思路是:1)将系统转化为输出预测器形式; 2)对于输出预测器,理想控制信号和策略效用函数可分别用行动网络和评论网络来逼近; 3)基于梯度下降法设计权值更新规则,构造最优控制信号。基于差分李雅普诺夫方法证明了系统的稳定性。最后,通过数值仿真验证了该方案的有效性.
In this paper, we propose an optimal control scheme-based adaptive neural network design for a class of unknown nonlinear discrete-time systems. The controlled systems are in a block-triangular multi-input-multi-output pure-feedback structure, i.e., there are both state and input couplings and nonaffine functions to be included in every equation of each subsystem. The design objective is to provide a control scheme, which not only guarantees the stability of the systems, but also achieves optimal control performance. The main contribution of this paper is that it is for the first time to achieve the optimal performance for such a class of systems. Owing to the interactions among subsystems, making an optimal control signal is a difficult task. The design ideas are that: 1) the systems are transformed into an output predictor form; 2) for the output predictor, the ideal control signal and the strategic utility function can be approximated by using an action network and a critic network, respectively; and 3) an optimal control signal is constructed with the weight update rules to be designed based on a gradient descent method. The stability of the systems can be proved based on the difference Lyapunov method. Finally, a numerical simulation is given to illustrate the performance of the proposed scheme.
使用单网络 ADP 的连续时间非线性系统非零和微分博弈的近最优控制
DOI: 10.1109/tsmcb.2012.2203336
发表时间: 2013-02
影响因子: 11.8
作者:
Huaguang Zhang;Lili Cui;Yanhong Luo
通讯作者: Yanhong Luo
基于有限逼近误差的离散时间迭代自适应动态规划
DOI: 10.1109/tcyb.2014.2354377
发表时间: 2014-09
影响因子: 11.8
作者:
Wei, Qinglai;Wang, Fei-Yue;Liu, Derong;Yang, Xiong
通讯作者: Yang, Xiong
DOI: 10.1049/iet-cta.2013.0472
发表时间: 2013-11
影响因子: 2.6
作者:
Xiong Yang;Derong Liu;Yuzhu Huang
通讯作者: Xiong Yang;Derong Liu;Yuzhu Huang
DOI: 10.1016/j.fss.2007.08.015
发表时间: 2008-04
期刊: Fuzzy Sets Syst.
影响因子: --
作者:
A. Boulkroune;M. Tadjine;M. M'Saad;M. Farza
通讯作者: A. Boulkroune;M. Tadjine;M. M'Saad;M. Farza
离散时间非线性系统的策略迭代自适应动态规划算法
DOI: 10.1109/tnnls.2013.2281663
发表时间: 2014-03-01
影响因子: 10.4
作者:
Liu, Derong;Wei, Qinglai
通讯作者: Wei, Qinglai