arrow
Return

Data-Driven Policy Iteration for Nonlinear Optimal Control Problems

delete2023-10-01
delete4
PRE
AI
C
Corrado Possieri *
M
Mario Sassano
DOI:10.1109/TNNLS.2022.3142501delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The design of optimal control laws for nonlinear systems is tackled without knowledge of the underlying plant and of a functional description of the cost function. The proposed data-driven method is based only on real-time measurements of the state of the plant and of the (instantaneous) value of the reward signal and relies on a combination of ideas borrowed from the theories of optimal and adaptive control problems. As a result, the architecture implements a policy iteration strategy in which, hinging on the use of neural networks, the policy evaluation step and the computation of the relevant information instrumental for the policy improvement step are performed in a purely continuous-time fashion. Furthermore, the desirable features of the design method, including convergence rate and robustness properties, are discussed. Finally, the theory is validated via two benchmark numerical simulations.
Keywords:
Optimal control
Costs
Neural networks
Real-time systems
Nonlinear dynamical systems
Closed loop systems
Learning systems
Data-driven methods
nonlinear systems
optimal control
policy iteration

Journal

IEEE Transactions on Neural Networks and Learning Systems cover
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
Papers:
7.6K
Citations:
7.2W

Organization

C
consiglio nazionale delle ricerche (cnr)
Scholars:
6.2W
Papers: 5.7W
Citations: 48
U
University of Rome Tor Vergata
Scholars:
2.5W
Papers: 1.8W
Citations: 2.0W
Cited Papers

Cited Papers

A novel actor-critic-identifier architecture for approximate optimal control of uncertain nonlinear systems
err2013-01-01
err481
PREAI
errBhasin, S.; Kamalapurkar, R.; Johnson, M.; Vamvoudakis, K. G.; Lewis, F. L.; Dixon, W. E.
errShare
errSave
Detecting Changes in Water Quality Data
err2008-01-01
err0
PREAI
errSean A. McKenna; Mark Wilson; Katherine A. Klise
errShare
errSave
Adaptive Optimal Control for Large-Scale Nonlinear Systems
err2017-11-01
err32
errOAAI
errMichailidis, Iakovos; Baldi, Simone; Kosmatopoulos, Elias B.; Ioannou, Petros A.
errShare
errSave
Moats and Drawbridges: An Isolation Primitive for Reconfigurable Hardware Based Systems
err2007-05-01
err0
errOAAI
errTed Huffmire; Brett Brotherton; Gang Wang; Timothy Sherwood; Ryan Kastner; Timothy Levin; Thuy Nguyen; Cynthia Irvine
errShare
errSave
researcher View more