返回
Data-driven distributed output consensus control for multi-agent systems with unknown internal state
DOI:10.1016/j.neucom.2024.128868.png)
摘要
En 中文
This paper discusses the topic of optimal output consensus control of linear time-invariant discrete-time multi-agent systems (MASs). The attainment of optimal output consensus control for MASs hinges upon the resolution of the interconnected Hamilton-Jacobi-Bellman equation, a task typically precluded by the inherent intractability of analytical solutions. Furthermore, most real-world systems are too complex to obtain the internal states of systems. To address these issues, a modified deep Q-learning network is constructed using current and historical system data rather than a precise model of the system. First, reconstructing the internal state of each agent using an adaptive distributed observer based on output feedback prevents the system instability brought on by the augmented systems. Then the local error system of the agent can be redefined. Based on the redefined error system, a data-driven adaptive dynamic programming (ADP) method is introduced, realized by using the actor-critic neural network structure. In addition, an experience replay strategy is proposed to reduce the propagation of estimation bias and improve the learning speed. Finally, the comparative numerical simulations substantiate the efficacy of the proposed algorithm in a quantifiable manner.
Keyword:
ECONOMIC-DISPATCH
OPTIMIZATION
SYNCHRONIZATION
NETWORKS
AGENTS
期刊
IF:
6.5
论文数:
2.5W
被引数:
6.5W
机构
引用论文
Adaptive Tracking Control for Perturbed Strict-Feedback Nonlinear Systems Based on Optimized Backstepping Technique基于优化反推技术的扰动严格反馈非线性系统自适应跟踪控制
Dreidimensionale Charakterisierung der Kornform und Scharfkantigkeit von Gesteinskörnungen mittels Röntgen-Computertomographie/3D characterisation of the grain sphericity and angularity with the aid of computed tomography
Bauingenieur
IF0
Constrained Event-Triggered H∞ Control Based on Adaptive Dynamic Programming With Concurrent Learning基于并行学习自适应动态规划的约束事件触发h ∞ 控制

