arrow
返回

Invariant Adaptive Dynamic Programming for Discrete-Time Optimal Control

delete2020-11-01
delete31
delete
OA
AI
Y
Yuanheng Zhu
D
Dongbin Zhao
H
Haibo He *
DOI:10.1109/TSMC.2019.2911900delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
For systems that can only be locally stabilized, control laws and their effective regions are both important. In this paper, invariant policy iteration is proposed to solve the optimal control of discrete-time systems. At each iteration, a given policy is evaluated in its invariantly admissible region, and a new policy and a new region are updated for the next iteration. Theoretical analysis shows the method is regionally convergent to the optimal value and the optimal policy. Combined with sum-of-squares polynomials, the method is able to achieve the near-optimal control of a class of discrete-time systems. An invariant adaptive dynamic programming algorithm is developed to extend the method to scenarios where system dynamics is not available. Online data are utilized to learn the near-optimal policy and the invariantly admissible region. Simulated experiments verify the effectiveness of our method.
Keyword:
Optimal control
Discrete-time systems
Heuristic algorithms
Dynamic programming
Convergence
Artificial intelligence
Nonlinear systems
Adaptive dynamic programming
discrete-time systems
invariant admissibility
optimal control
policy iteration
sum of squares
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Cybernetics 封面图
IEEE Transactions on Cybernetics
IF:
10.5
论文数:
1.1W
被引数:
5.0W

机构

C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

err分享
err收藏
A brief summary of hypoxia on the northern Gulf of Mexico continental shelf: 1985–1988
err1991-12-01
err0
PREAI
errNancy N. Rabalais; R. Eugene Turner; William J. Wiseman; Donald F. Boesch
err分享
err收藏
学者 查看更多内容