arrow
Return

Asynchronous multi-agent deep reinforcement learning under partial observability

delete2025-02-06
delete0
PRE
AI
Y
Yuchen Xiao *
W
Weihao Tan
J
Joshua Hoffman
T
Tian Xia
C
Christopher Amato
DOI:10.1177/02783649241306124delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The state-of-the-art multi-agent reinforcement learning (MARL) methods provide promising solutions to a variety of complex problems. Yet, these methods all assume that agents perform primitive actions in a synchronized manner, making them impractical for long-horizon real-world multi-robot tasks that inherently require robots to asynchronously reason about action selection at varying time durations. To solve this problem, we first propose a group of value-based cooperative MARL approaches for asynchronous execution using temporally extended macro-actions. Here, agents perform asynchronous learning and decision-making with macro-action-value functions in three paradigms: decentralized learning and control, centralized learning and control, and centralized training for decentralized execution (CTDE). Building on the above work, we formulate a set of macro-action-based policy gradient algorithms under the three training paradigms, where agents directly optimize their parameterized policies in an asynchronous manner. We evaluate our methods both in simulation and on real robots over a variety of realistic domains. Empirical results demonstrate the effectiveness of our algorithms for learning high-quality and asynchronous solutions with macro-actions in large multi-agent problems that were previously unsolvable via primitive-action-based approaches. The proposed approaches represent the first general MARL methods for temporally extended actions and serve as the foundation for future methods in the area.
Keywords:
Multi-agent
reinforcement learning
macro-actions

Journal

International Journal of Robotics Research cover
International Journal of Robotics Research
IF:
5
Papers:
2.4K
Citations:
1.5W

Organization

No organization information available
Cited Papers

Cited Papers

Multi-agent Double Deep Q-Networks
err2017-08-09
err0
PREAI
errDavid Simões; Nuno Lau; Luís Paulo Reis
errShare
errSave
Policy Search for Multi-Robot Coordination under Uncertainty
err2015-07-13
err0
errOAAI
errChristopher Amato; George Konidaris; Ariel Anders; Gabriel Cruz; Jonathan How; Leslie Kaelbling
errShare
errSave
Counterfactual Multi-Agent Policy Gradients
err2018-04-29
err0
errOAAI
errJakob Foerster; Gregory Farquhar; Triantafyllos Afouras; Nantas Nardelli; Shimon Whiteson
errShare
errSave
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation
err2014-01-01
err0
errOAAI
errKyunghyun Cho; Bart van Merrienboer; Caglar Gulcehre; Dzmitry Bahdanau; Fethi Bougares; Holger Schwenk; Yoshua Bengio
errShare
errSave
errShare
errSave
errShare
errSave
errShare
errSave
Online Planning for Target Object Search in Clutter under Partial Observability
err2019-05-01
err0
PREAI
errYuchen Xiao; Sammie Katt; Andreas ten Pas; Shengjian Chen; Christopher Amato
errShare
errSave
researcher View more