返回
Automatic Curriculum Learning for Large-Scale Cooperative Multiagent Systems
DOI:10.1109/TETCI.2022.3209655.png)
摘要
En 中文
Recently, a lot of works have been devoted to researching how agents can learn efficient cooperation in multiagent systems. However, it still remains challenging in large-scale multiagent systems (MASs) due to the complex dynamics between the agents and environment and the dimension explosion of state-action space. In this paper, we propose a novel MultiAgent Automatic Curriculum Learning method (MA-ACL) to solve learning problems of large-scale cooperative MASs by beginning from learning on a multiagent scenario with a few agents and automatically progressively increasing the number of agents. An evaluation mechanism based on self-supervised learning is innovatively designed to automatically generate appropriate curricula with a progressively increasing number of agents. Moreover, since the observation state dimension of agents varies across curricula and the learned policy knowledge needs to be effectively encoded, we design a new Distributed Transferable Relation-modeling Policy network structure (DTRP) to handle the dynamic size of the network input and model relational knowledge between agents and their surrounding environment. Simulation results show that the proposed MA-ACL using DTRP can significantly improve the performance of large-scale multiagent learning compared with manual or non curriculum learning methods, and DTRP greatly boosts the performance of MA-ACL.
Keyword:
Task analysis
Games
Markov processes
Training
Multi-agent systems
Computational intelligence
Observability
Automatic curriculum learning
large-scale multiagent systems
multiagent reinforcement learning
期刊
I
IF:
6.5
论文数:
1.4K
被引数:
4.5K
机构
引用论文
Parallel amygdala and inferotemporal activation reflect emotional intensity and fear relevance
NeuroImage
IF0
Foam-formed cellulose composite materials with potential applications in sound insulation泡沫形成的纤维素复合材料在隔音领域的潜在应用

