arrow
返回

Asynchronous iterative Q-learning based tracking control for nonlinear discrete-time multi-agent systems

delete2024-12-01
delete1
PRE
AI
T
Tao Dong *
T
Tingwen Huang
DOI:10.1016/j.neunet.2024.106667delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This paper addresses the tracking control problem of nonlinear discrete-time multi-agent systems (MASs). First, a local neighborhood error system (LNES) is constructed. Then, a novel tracking algorithm based on asynchronous iterative Q-learning (AIQL) is developed, which can transform the tracking problem into the optimal regulation of LNES. The AIQL-based algorithm has two Q values Q(i)(A) and Q(i)(B) for each agent i , where Q(i)(A) is used for improving the control policy and Q(i)(B) is used for evaluating the value of the control policy. Moreover, the convergence of LNES is given. It is shown that the LNES converges to 0 and the tracking problem is solved. A neural network-based actor-critic framework is used to implement AIQL. The critic network of AIQL is composed of two neural networks, which are used for approximating Q(i)(A) and Q(i)(B) respectively. Finally, simulation results are given to verify the performance of the developed algorithm. It is shown that the AIQLbased tracking algorithm has a lower cost value and faster convergence speed than the IQL-based tracking algorithm.
Keyword:
Multi-agent
Discrete-time
Asynchronous iterative Q-learning
Tracking control

期刊

Neural Networks 封面图
Neural Networks
IF:
6.3
论文数:
7.9K
被引数:
3.0W

机构

S
southwest university - china
学者数:
2.6W
论文数: 1.9W
被引数: 21
S
Shenzhen University of Advanced Technology
学者数:
339
论文数: 330
被引数: 1
引用论文

引用论文

A review of the diagnosability of control systems with applications to spacecraft
err2020-01-01
err29
PREAI
errWang, Dayi; Fu, Fangzhou; Li, Wenbo; Tu, Yuanyuan; Liu, Chengrui; Liu, Wenjing
err分享
err收藏
Adaptive memory-based event-triggering resilient LFC for power system under DoS attack
err2023-08-01
err13
PREAI
errLiu, Xingyue; Shi, Kaibo; Cheng, Jun; Wen, Shiping; Liu, Yajuan
err分享
err收藏
err分享
err收藏
Finite-time consensus control for multi-agent systems with full-state constraints and actuator failures
err2023-01-01
err40
PREAI
errWang, Jianhui; Yan, Yancheng; Liu, Zhi; Chen, C. L. Philip; Zhang, Chunliang; Chen, Kairui
err分享
err收藏
学者 查看更多内容