arrow
返回

Q-Learning-Based Model Predictive Control for Nonlinear Continuous-Time Systems

delete2020-09-11
delete23
PRE
AI
H
Hao Zhang
S
Shaoyuan Li
Y
Yi Zheng *
DOI:10.1021/acs.iecr.0c02321delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this paper, a Q-learning-based model predictive control using the Lyapunov technique (Q-LMPC) is proposed for the control of a class of continuous nonlinear systems with complicated dynamics, whose accurate mathematical model is hard or more efforts are needed to be obtained. The proposed method learns an approximate control policy of Lyapunov-based model predictive control (MPC) based on data, and the learned control policy is approximated by a neural network, called actor network, which guarantees the computational efficiency of control law regardless of the complexity of system dynamics. In the proposed MPC, a finite-horizon iterative reinforcement learning (RL) algorithm is developed to obtain the closed-loop optimal/suboptimal solutions of a nonlinear optimization problem with Lyapunov constraints in the idle time of the controller. In the meantime, the critic and actor neural networks used for the implementation of MPC are updated in an iterative manner according to these solutions. The convergence of the iterative Q-LMPC method and the Lyapunov stability of the closed-loop control system are analyzed. The simulation result shows the effectiveness of the proposed Q-LMPC method.
Keyword:
STABILIZATION
STATE
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

I
Industrial and Engineering Chemistry Research
IF:
3.9
论文数:
4.0W
被引数:
9.6W

机构

S
shanghai jiao tong university
学者数:
15.7W
论文数: 11.7W
被引数: 159
引用论文

引用论文

err分享
err收藏
Reinforcement Learning for Port-Hamiltonian Systems
err2015-05-01
err31
errOAAI
errSprangers, Olivier; Babuska, Robert; Nageshrao, Subramanya P.; Lopes, Gabriel A. D.
err分享
err收藏
Constructive nonlinear control: a historical perspective
err2001-05-01
err647
PREAI
errKokotovic, P; Arcak, M
err分享
err收藏
学者 查看更多内容