arrow
返回

Coordinating Multi-Agent Reinforcement Learning via Dual Collaborative Constraints

delete2025-02-01
delete0
PRE
AI
C
Chao Li
S
Shaokang Dong
S
Shangdong Yang
Y
Yujing Hu
W
Wenbin Li
高扬 封面图
高扬 (Yang Gao) *
DOI:10.1016/j.neunet.2024.106858delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Many real-world multi-agent tasks exhibit a nearly decomposable structure, where interactions among agents within the same interaction set are strong while interactions between different sets are relatively weak. Efficiently modeling the nearly decomposable structure and leveraging it to coordinate agents can enhance the learning efficiency of multi-agent reinforcement learning algorithms for cooperative tasks, while existing works typically fail. To overcome this limitation, this paper proposes a novel algorithm named Dual Collaborative Constraints (DCC) that identifies the interaction sets as subtasks and achieves both intra-subtask and intersubtask coordination. Specifically, DCC employs a bi-level structure to periodically distribute agents into multiple subtasks, and proposes both local and global collaborative constraints based on mutual information to facilitate both intra-subtask and inter-subtask coordination among agents. These two constraints ensure that agents within the same subtask reach a consensus on their local action selections and all of them select superior joint actions that maximize the overall task performance. Experimentally, we evaluate DCC on various cooperative multi-agent tasks, and its superior performance against multiple state-of-the-art baselines demonstrates its effectiveness.
Keyword:
Multi-agent reinforcement learning
Cooperative tasks
Coordination
Nearly decomposable structure

期刊

Neural Networks 封面图
Neural Networks
IF:
6.3
论文数:
7.8K
被引数:
3.0W

机构

N
nanjing university
学者数:
7.8W
论文数: 5.6W
被引数: 87
引用论文

引用论文

Priming
err2002-01-01
err0
PREAI
errAnthony D. Wagner; Wilma Koutstaal
err分享
err收藏
Risk of second primary papillary thyroid cancer among adult cancer survivors in the United States, 2000-2015
err2020-02-01
err0
errOAAI
errSara J. Schonfeld; Lindsay M. Morton; Amy Berrington de González; Rochelle E. Curtis; Cari M. Kitahara
err分享
err收藏
err分享
err收藏
Deep Reinforcement Learning With Modulated Hebbian Plus Q-Network Architecture
err2022-05-01
err13
errOAAI
errLadosz, Pawel; Ben-Iwhiwhu, Eseoghene; Dick, Jeffery; Ketz, Nicholas; Kolouri, Soheil; Krichmar, Jeffrey L.; Pilly, Praveen K.; Soltoggio, Andrea
err分享
err收藏
Strategic Densification With UAV-BSs in Cellular Networks
err2018-06-01
err0
PREAI
errFaraj Lagum; Irem Bor-Yaliniz; Halim Yanikomeroglu
err分享
err收藏
学者 查看更多内容