返回
Evolution-Guided Adaptive Dynamic Programming for Nonlinear Optimal Control
DOI:10.1109/TSMC.2024.3417230.png)
摘要
En 中文
In this article, an evolution-guided adaptive dynamic programming (EGADP) algorithm is developed to address the optimal regulation problems for the nonlinear systems. In the traditional adaptive dynamic programming algorithms, policy improvement is typically reliant on the gradient information, according to the first order necessity condition. However, these methods encounter limitations when calculating the gradient information becomes infeasible or system dynamics is not differentiable. In response to this challenge, the evolutionary computation is harnessed by EGADP to search for a superior policy during policy improvement. Therefore, compared with the traditional methods, scenarios that gradient information is unavailable can effectively be handled by EGADP. Additionally, the convergence of the algorithm is proven to enhance the rigorousness of the developed method. Finally, the three simulation experiments with realistic physical backgrounds are conducted to comprehensively demonstrate the effectiveness of the established method from different perspectives.
Keyword:
Adaptive critic designs
adaptive dynamic programming (ADP)
evolutionary computation (EC)
intelligent control
optimal control
reinforcement learning (RL)
期刊
IF:
10.5
论文数:
1.1W
被引数:
5.0W
机构
引用论文
Value iteration and adaptive dynamic programming for data-driven adaptive optimal control design数据驱动的自适应最优控制设计的值迭代和自适应动态规划
AUTOMATICA
IF5.9
Advanced value iteration for discrete-time intelligent critic control: A survey离散时间智能评论家控制的高级值迭代: 一项调查
Ac-Electrogravimetry Study of Electroactive Thin Films. II. Application to Polypyrrole电活性薄膜的交流电重分析研究。二。聚吡咯的应用
Study of bi-directional buck-boost converter topologies for application in electrical vehicle motor drives应用于电动汽车电机驱动的双向buck-boost变换器拓扑研究
Discrete-Time Stable Generalized Self-Learning Optimal Control With Approximation Errors具有逼近误差的离散稳定广义自学习最优控制
System Stability of Learning-Based Linear Optimal Control With General Discounted Value Iteration具有一般折现值迭代的基于学习的线性最优控制的系统稳定性

