arrow
返回

Adaptive Dynamic Programming for Nonlinear-Constrained H8 Control

delete2023-07-01
delete13
PRE
AI
杨雄 (Xiong Yang) *
M
Mengmeng Xu
Q
Qinglai Wei
DOI:10.1109/TSMC.2023.3247888delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This article considers the H-infinity control problem of nonlinear systems having unavailable dynamics and asymmetric saturating actuators. Initially, such an H-infinity control problem is converted into the zero-sum game with a nonquadratic cost function being introduced. Then, in order to solve the Hamilton-Jacobi-Isaacs equation arising in the zero-sum game, a simultaneous policy iteration (SPI) algorithm is developed under the adaptive dynamic programming framework. Meanwhile, it is proved that the convergence of the SPI algorithm in essence amounts to the convergence of the sequential PI algorithm. To implement the SPI algorithm, the critic, the actor, and the perturbation neural networks (NNs) are, respectively, constructed to estimate the cost function, the control policy, and the perturbation. The three NNs' weights are simultaneously determined by using the least-squares method together with the Monte Carlo integration technique. A remarkable characteristic of such an SPI algorithm is that arbitrary control policies and perturbations are applicable in the learning process. This makes system's information be able to be replaced by the data collected along system's trajectories in advance. More importantly, the persistence of the excitation condition is not required. Finally, simulations of two nonlinear examples are given to validate the present SPI algorithm.
Keyword:
Perturbation methods
Games
Game theory
Nonlinear systems
Heuristic algorithms
Artificial neural networks
Dynamic programming
Adaptive
approximate dynamic programming (ADP)
asymmetric constraint
neural network (NN) control
simultaneous policy iteration (SPI)

期刊

IEEE Transactions on Cybernetics 封面图
IEEE Transactions on Cybernetics
IF:
10.5
论文数:
1.1W
被引数:
5.0W

机构

T
tianjin university
学者数:
8.0W
论文数: 5.8W
被引数: 88
C
chinese academy of sciences
学者数:
56.7W
论文数: 44.9W
被引数: 704
引用论文

引用论文

Event-driven H∞ control with critic learning for nonlinear systems
err2020-12-01
err14
PREAI
errYang, Xiong; Gao, Zhongke; Zhang, Jinhui
err分享
err收藏
学者 查看更多内容