返回
Data-Driven Policy Iteration for Nonlinear Optimal Control Problems
DOI:10.1109/TNNLS.2022.3142501.png)
摘要
En 中文
The design of optimal control laws for nonlinear systems is tackled without knowledge of the underlying plant and of a functional description of the cost function. The proposed data-driven method is based only on real-time measurements of the state of the plant and of the (instantaneous) value of the reward signal and relies on a combination of ideas borrowed from the theories of optimal and adaptive control problems. As a result, the architecture implements a policy iteration strategy in which, hinging on the use of neural networks, the policy evaluation step and the computation of the relevant information instrumental for the policy improvement step are performed in a purely continuous-time fashion. Furthermore, the desirable features of the design method, including convergence rate and robustness properties, are discussed. Finally, the theory is validated via two benchmark numerical simulations.
Keyword:
Optimal control
Costs
Neural networks
Real-time systems
Nonlinear dynamical systems
Closed loop systems
Learning systems
Data-driven methods
nonlinear systems
optimal control
policy iteration
期刊
IF:
8.9
论文数:
7.5K
被引数:
7.2W
机构
引用论文
Adaptive Neural Networks Finite-Time Optimal Control for a Class of Nonlinear Systems一类非线性系统的自适应神经网络有限时间最优控制
A novel actor-critic-identifier architecture for approximate optimal control of uncertain nonlinear systems一种用于不确定非线性系统的近似最优控制的新型参与者-批评者-标识符体系结构
AUTOMATICA
IF5.9

