arrow
返回

Value Iteration for Continuous-Time Linear Time-Invariant Systems

delete2023-05-01
delete3
PRE
AI
C
Corrado Possieri *
M
Mario Sassano
DOI:10.1109/TAC.2022.3169688delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Two data-driven strategies for value iteration in linear quadratic optimal control problems over an infinite horizon are proposed. The two architectures share common features, since they both consist of a purely continuous-time control architecture and are based on the forward integration of the differential Riccati equation (DRE). They profoundly differ, instead, in the estimation mechanism of the vector field of the underlying DRE from collected data: The first relies on a characterization of properties of the advantage function associated to the problem, whereas the second is inspired by tools from adaptive control theory and ensures semi-global exponential convergence to the optimal solution. Advantages and drawbacks of the architectures are discussed, while the performance is validated via a benchmark numerical example.
Keyword:
Optimal control
Costs
Riccati equations
Reinforcement learning
Convergence
Adaptive control
Trajectory
linear systems
optimal control
reinforcement learning

期刊

IEEE Transactions on Automatic Control 封面图
IEEE Transactions on Automatic Control
IF:
7
论文数:
1.3W
被引数:
6.7W

机构

U
University of Rome Tor Vergata
学者数:
2.5W
论文数: 1.8W
被引数: 2.0W
引用论文

引用论文

err分享
err收藏
Cation Complexation, Photochromism, and Aggregation of Copolymers Carrying Crown Ether and Spirobenzopyran Moieties at the Side Chains
err2003-01-15
err0
PREAI
errKimura Keiichi; Makoto Nakamura; Hidefumi Sakamoto; Ryoko Mizutani Uda; Miwa Sumida; Masaaki Yokoyama
err分享
err收藏
Inhibitory effects of E-5110 on interleukin-1 generation from human monocytes
err1989-06-01
err0
PREAI
errH. Shirota; M. Goto; R. Hashida; I. Yamatsu; K. Katayama
err分享
err收藏
没有更多内容