arrow
返回

Data-Based Optimal Consensus Control for Multiagent Systems With Policy Gradient Reinforcement Learning

delete2022-08-01
delete31
PRE
AI
X
Xindi Yang
张浩 封面图
张浩 (Hao Zhang)
Z
Zhuping Wang *
DOI:10.1109/TNNLS.2021.3054685delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This article investigates the optimally distributed consensus control problem for discrete-time multiagent systems with completely unknown dynamics and computational ability differences. The problem can be viewed as solving nonzero-sum games with distributed reinforcement learning (RL), and each agent is a player in these games. First, to guarantee the real-time performance of learning algorithms, a data-based distributed control algorithm is proposed for multiagent systems using offline system interaction data sets. By utilizing the interactive data produced during the run of a real-time system, the proposed algorithm improves system performance based on distributed policy gradient RL. The convergence and stability are guaranteed based on functional analysis and the Lyapunov method. Second, to address asynchronous learning caused by computational ability differences in multiagent systems, the proposed algorithm is extended to an asynchronous version in which executing policy improvement or not of each agent is independent of its neighbors. Furthermore, an actor-critic structure, which contains two neural networks, is developed to implement the proposed algorithm in synchronous and asynchronous cases. Based on the method of weighted residuals, the convergence and optimality of the neural networks are guaranteed by proving the approximation errors converge to zero. Finally, simulations are conducted to show the effectiveness of the proposed algorithm.
Keyword:
Multi-agent systems
Consensus control
Games
Heuristic algorithms
Dynamic programming
Synchronization
Reinforcement learning
Asynchronous learning
data-based control
nonzero-sum games
optimal distributed consensus control
policy gradient (PG) reinforcement learning (RL)

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

T
tongji university
学者数:
7.9W
论文数: 6.0W
被引数: 98
引用论文

引用论文

Development of ceramic porcelain stoneware pastes by the revalorization of Colombian clays subjected to bleaching process
err2019-09-01
err0
PREAI
errYudi E. Ramírez Calderón; Carlos A. Nieto Rangel; Jorge Llop Pla; Ester Barrachina Albert; Jesús S. Valencia Rios; Juan B. Carda Castelló
err分享
err收藏
err分享
err收藏
Data-Driven Distributed Optimal Consensus Control for Unknown Multiagent Systems With Input-Delay
err2019-06-01
err84
PREAI
errZhang, Huaipin; Yue, Dong; Dou, Chunxia; Zhao, Wei; Xie, Xiangpeng
err分享
err收藏
学者 查看更多内容