返回
Stability analysis of heuristic dynamic programming algorithm for nonlinear systems
DOI:10.1016/j.neucom.2014.08.046.png)
摘要
En 中文
In this paper, a value-iteration based heuristic dynamic programming (HDP) algorithm is developed to solve the optimal control for the continuous time affine nonlinear systems. First, a rigorous convergence proof of the HDP algorithm is given. Second, stability issues of the HDP algorithm for nonlinear systems are investigated. It is commonly believed that the main drawback of the HDP algorithm is that only the limit function of the iterative control sequence is proved to be stabilized, thus infinite iterations are executed. To confront this problem, we present a novel stability result for the HDP algorithm, which indicates that the resulting iterative control laws after finite iterations can guarantee the closed-loop stability. A similar stability result is also obtained for the discrete time nonlinear systems. Therefore, the practicality of the HDP algorithm is greatly improved. Single neural network (NN) structure is employed to implement the algorithm. It should be pointed that the algorithm can be implemented without knowing the internal dynamics of the systems. Finally, two numerical examples are given to demonstrate the effectiveness of the developed methods. (c) 2014 Elsevier B.V. All rights reserved.
Keyword:
Convergence
Stability
Heuristic dynamic programming (HDP)
Optimal control
Value-iteration
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.5
论文数:
2.5W
被引数:
6.5W
机构
引用论文
Reactive power control of grid-connected wind farm based on adaptive dynamic programming
NEUROCOMPUTING
IF6.5
Optimal control of unknown nonaffine nonlinear discrete-time systems based on adaptive dynamic programming基于自适应动态规划的未知非仿射非线性离散系统最优控制
AUTOMATICA
IF5.9
An iterative adaptive dynamic programming method for solving a class of nonlinear zero-sum differential games求解一类非线性零和微分对策的迭代自适应动态规划方法
AUTOMATICA
IF5.9

