arrow
返回

Reinforcement Learning Controller Design for Discrete-Time-Constrained Nonlinear Systems With Weight Initialization Method

delete2024-04-01
delete2
PRE
AI
J
Jiahui Xu
J
Jingcheng Wang *
Y
Yanjiu Zhong
J
Jun Rao
S
Shunyu Wu
DOI:10.1109/TSMC.2023.3344883delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Extensive research has been dedicated to reinforcement learning (RL) for acquiring proficient optimal controllers through interactions with the environment. However, real-world demands, including enhanced safety performance, introduce considerable challenges to the present design of optimal controllers rooted in RL algorithms. A novel approach is introduced in this article for designing RL-based optimal controllers, employing a control barrier function (CBF) alongside a nonquadratic loss function related to the control signal. The aim is to enable the agent to learn the optimal controller in a secure and efficient manner. To tackle the instability issue in neural network training inherent to traditional RL-based controller design processes, the nonlinear model predictive control (NMPC) technique is employed for initializing the controller network's weights. A formal demonstration of the method's optimality is presented. Numerical simulations validate the proposed approach, illustrating its capacity to effectively learn the optimal controller while adhering to the input and state constraints of the system.
Keyword:
Control barrier function (CBF)
nonlinearmodel predictive control (NMPC)
nonlinear systems
reinforce-ment learning (RL)
state constraints

期刊

IEEE Transactions on Cybernetics 封面图
IEEE Transactions on Cybernetics
IF:
10.5
论文数:
1.1W
被引数:
5.0W

机构

S
shanghai jiao tong university
学者数:
15.6W
论文数: 11.6W
被引数: 159