arrow
返回

Stabilizing reinforcement learning control: A modular framework for optimizing over all stable behavior

delete2024-06-01
delete0
delete
OA
AI
N
Nathan P. Lawrence
P
Philip D. Loewen *
S
Shuyuan Wang
M
Michael G. Forbes
R
R. Bhushan Gopaluni
DOI:10.1016/j.automatica.2024.111642delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
We propose a framework for the design of feedback controllers that combines the optimization -driven and model -free advantages of deep reinforcement learning with the stability guarantees provided by using the Youla-Ku & ccaron;era parameterization to define the search domain. Recent advances in behavioral systems allow us to construct a data -driven internal model; this enables an alternative realization of the Youla-Ku & ccaron;era parameterization based entirely on input-output exploration data. Perhaps of independent interest, we formulate and analyze the stability of such data -driven models in the presence of noise. The Youla-Ku & ccaron;era approach requires a stable parameterfor controller design. For the training of reinforcement learning agents, the set of all stable linear operators is given explicitly through a matrix factorization approach. Moreover, a nonlinear extension is given using a neural network to express a parameterized set of stable operators, which enables seamless integration with standard deep learning libraries. Finally, we show how these ideas can also be applied to tune fixed -structure controllers (c) 2024 The Author(s). Published by Elsevier Ltd. This is an open access article under the CC BY license (http://creativecommons.org/licenses/by/4.0/).
Keyword:
Reinforcement learning
Data -driven control
Youla-Ku & ccaron
era parameterization
Neural networks
Stability
Process control
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Automatica 封面图
Automatica
IF:
5.9
论文数:
1.2W
被引数:
5.2W

机构

H
honeywell
学者数:
395
论文数: 280
被引数: 0
U
University of British Columbia
学者数:
7.0W
论文数: 6.1W
被引数: 8.6W
引用论文

引用论文

Robust reinforcement learning control with static and dynamic stability
err2002-01-07
err49
errOAAI
errKretchmar, RM; Young, PM; Anderson, CW; Hittle, DC; Anderson, ML; Delnero, CC
err分享
err收藏
err分享
err收藏
Deep reinforcement learning with shallow controllers: An experimental application to PID tuning基于浅层控制器的深度强化学习: PID整定的实验应用
err2022-04-01
err67
errOAAI
errLawrence, Nathan P.; Forbes, Michael G.; Loewen, Philip D.; McClement, Daniel G.; Backstrom, Johan U.; Gopaluni, R. Bhushan
err分享
err收藏
Density and viscosity of three (2,2,2-trifluoroethanol + 1-butyl-3-methylimidazolium) ionic liquid binary systems
err2014-03-01
err0
PREAI
errJosefa Salgado; Teresa Regueira; Luis Lugo; Javier Vijande; Josefa Fernández; Josefa García
err分享
err收藏
Cover Feature: Electrochemical Carbon Dioxide Splitting (ChemElectroChem 6/2019)
err2019-02-27
err0
errOAAI
errJiafang Xie; Yiyin Huang; Maoxiang Wu; Yaobing Wang
err分享
err收藏
学者 查看更多内容