返回
Generative subgoal oriented multi-agent reinforcement learning through potential field
DOI:10.1016/j.neunet.2024.106552.png)
摘要
En 中文
Multi-agent reinforcement learning (MARL) effectively improves the learning speed of agents in sparse reward tasks with the guide of subgoals. However, existing works sever the consistency of the learning objectives of the subgoal generation and subgoal reached stages, thereby significantly inhibiting the effectiveness of subgoal learning. To address this problem, we propose a novel Potential field Subgoal-based Multi-Agent reinforcement learning (PSMA) method, which introduces the potential field (PF) to unify the two-stage learning objectives. Specifically, we design a state-to-PF representation model that describes agents' states as potential fields, allowing easy measurement of the interaction effect for both allied and enemy agents. With the PF representation, a subgoal selector is designed to automatically generate multiple subgoals for each agent, drawn from the experience replay buffer that contains both individual and total PF values. Based on the determined subgoals, we define an intrinsic reward function to guide the agent to reach their respective subgoals while maximizing the joint action-value. Experimental results show that our method outperforms the state-of-the-art MARL method on both StarCraft II micro-management (SMAC) and Google Research Football (GRF) tasks with sparse reward settings.
Keyword:
Multi-agent reinforcement learning
Subgoal generation
Potential field
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.3
论文数:
8.2K
被引数:
3.0W
机构
暂无机构信息
引用论文
Chirality sensing of various biomolecules with helical poly(phenylacetylene)s bearing acidic functional groups in water带有酸性官能团的螺旋聚 (苯乙炔) 在水中对各种生物分子的手性传感
Z -2-Phenyl-4-[(S)- 2,2-dimethyl-1,3-dioxolan-4-ylmethylen]-5(4H)-oxazolone as the Dienophile in Asymmetric Diels-Alder Reactions
Tetrahedron
IF0
Effect of calcium on the cryopreservation of Lactobacillus bulgaricus in different freezing media
Cryobiology
IF0
Managing engineering systems with large state and action spaces through deep reinforcement learning通过深度强化学习管理具有大型状态和动作空间的工程系统
没有更多内容

