arrow
返回

Accelerated Reinforcement Learning via Dynamic Mode Decomposition

delete2023-12-01
delete7
PRE
AI
V
Vrushabh S. Donge *
B
Bosen Lian
F
Frank L. Lewis
A
Ali Davoudi
DOI:10.1109/TCNS.2023.3259060delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This work applies the decomposition principle to discrete-time reinforcement learning (RL) to solve the optimal control problems for a network of subsystems. The control design is defined as a linear quadratic regulator graphical problem, where the performance function couples the subsystems' dynamics. We first present a model-free discrete-time RL algorithm based on online behaviors without using system dynamics. This could become a prohibitively long learning process for larger networks. To remedy this issue, we develop an efficient model-free RL algorithm based on dynamic mode decomposition. This decomposition method reduces the size of the measured data while the dynamic information of the original network is still retained. This algorithm is then implemented online. The proposed methodology is validated using examples of a consensus network and a power system network.
Keyword:
Dynamic mode decomposition
large-scale systems
optimal control
reinforcement learning (RL)

期刊

IEEE Transactions on Control of Network Systems 封面图
IEEE Transactions on Control of Network Systems
IF:
5
论文数:
1.6K
被引数:
5.8K

机构

U
university of texas system
学者数:
18.5W
论文数: 15.6W
被引数: 210