arrow
返回

Data-Driven Distributed Output Consensus Control for Partially Observable Multiagent Systems

delete2019-03-01
delete56
delete
OA
AI
H
He Jiang
H
Haibo He *
DOI:10.1109/TCYB.2017.2788819delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
This paper is concerned with a class of optimal output consensus control problems for discrete linear multiagent systems with the partially observable system state. Since the optimal control policy depends on the full system state which is not accessible for a partially observable system, traditionally, distributed observers are employed to recover the system state. However, in many situations, the accurate model of a real-world dynamical system might be difficult to obtain, which makes the observer design infeasible. Furthermore, the optimal consensus control policy cannot he analytically solved without system functions. To overcome these challenges, we propose a data-driven adaptive dynamic programming approach that does not require the complete system inner state. The key idea is to use the input and output sequence as an equivalent representation of the underlying state. Based on this representation, an adaptive dynamic programming algorithm is developed to generate the optimal control policy. For the implementation of this algorithm, we design a neural network-based actor-critic structure to approximate the local performance indices and the control polices. Two numerical simulations are used to demonstrate the effectiveness of our method.
Keyword:
Adaptive dynamical programming
consensus control
multiagent system
observability
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Cybernetics 封面图
IEEE Transactions on Cybernetics
IF:
10.5
论文数:
1.1W
被引数:
5.0W

机构

U
University of Rhode Island
学者数:
5.0K
论文数: 4.5K
被引数: 6.3K
引用论文

引用论文

err分享
err收藏
Ageing and degradation of paint film media
err1973-03-01
err0
PREAI
errEmil Krejcar; Otakar Kolář
err分享
err收藏
err分享
err收藏
Optimal model-free output synchronization of heterogeneous systems using off-policy reinforcement learning
err2016-09-01
err148
errOAAI
errModares, Hamidreza; Nageshrao, Subramanya P.; Lopes, Gabriel A. Delgado; Babuska, Robert; Lewis, Frank L.
err分享
err收藏
学者 查看更多内容