arrow
返回

Dynamic Deep Factor Graph for Multi-Agent Reinforcement Learning

delete2025-11-19
delete0
PRE
AI
Y
Yuchen Shi
段世红 (Shihong Duan)
C
Cheng Xu
R
Ran Wang
F
Fangwen Ye
C
Chau Yuen
DOI:10.1109/TPAMI.2025.3634378delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Multi-agent reinforcement learning (MARL) requires effective coordination among multiple decision-making agents to achieve joint goals. Approaches based on a global value function face the curse of dimensionality, while fully decomposed centralized training with decentralized execution (CTDE) methods often suffer from relative overgeneralization. Coordination graphs mitigate this issue but typically fail to capture dynamic collaboration patterns that evolve over time and across tasks. We propose Dynamic Deep Factor Graphs (DDFG), a value decomposition algorithm that represents the global value via factor graphs and learns graph structures on the fly through a graph-generation policy, adapting to evolving inter-agent relations. We provide a theoretical upper bound on the approximation error of high-order decompositions and reveal how the maximum order $D$ trades off accuracy against computation, offering guidance for balancing performance and cost. Using max-sum for inference, DDFG efficiently derives joint policies. Experiments on higher-order predator–prey and SMAC show consistent gains over strong value-decomposition baselines, demonstrating improved sample efficiency and robustness in complex settings.
Keyword:
Dynamic graph
factor graph
multi-agent reinforcement learning (MARL)
relative overgeneralization
dynamic collaboration

期刊

IEEE Transactions on Pattern Analysis and Machine Intelligence 封面图
IEEE Transactions on Pattern Analysis and Machine Intelligence
IF:
18.6
论文数:
864
被引数:
9.8W

机构

N
nanyang technological university
学者数:
2.5K
论文数: 1.6K
被引数: 1
U
university of science and technology beijing
学者数:
1.3W
论文数: 4.5K
被引数: 2