返回
Deep Multi-Agent Reinforcement Learning for Decentralized Active Hypothesis Testing
DOI:10.1109/ACCESS.2024.3430392.png)
摘要
En 中文
We consider a decentralized formulation of the active hypothesis testing (AHT) problem, where multiple agents gather noisy observations from the environment with the purpose of identifying the correct hypothesis. At each time step, agents have the option to select a sampling action. These different actions result in observations drawn from various distributions, each associated with a specific hypothesis. The agents collaborate to accomplish the task, where message exchanges between agents are allowed over a rate-limited communications channel. The objective is to devise a multi-agent policy that minimizes the Bayes risk. This risk comprises both the cost of sampling and the joint terminal cost incurred by the agents upon making a hypothesis declaration. Deriving optimal structured policies for AHT problems is generally mathematically intractable, even in the context of a single agent. As a result, recent efforts have turned to deep learning methodologies to address these problems, which have exhibited significant success in single-agent learning scenarios. In this paper, we tackle the multi-agent AHT formulation by introducing a novel algorithm rooted in the framework of deep multi-agent reinforcement learning. This algorithm, named Multi-Agent Reinforcement Learning for AHT (MARLA), operates at each time step by having each agent map its state to an action (sampling rule or stopping rule) using a trained deep neural network with the goal of minimizing the Bayes risk. We present a comprehensive set of experimental results that effectively showcase the agents' ability to learn collaborative strategies and enhance performance using MARLA. Furthermore, we demonstrate the superiority of MARLA over single-agent learning approaches. Finally, we provide an open-source implementation of the MARLA framework, for the benefit of researchers and developers in related domains.
Keyword:
Costs
Testing
Bayes methods
Collaboration
Task analysis
Software algorithms
Noise measurement
Deep reinforcement learning
Multi-agent systems
Active hypothesis testing (AHT)
controlled sensing for multihypothesis testing
decentralized inference
deep reinforcement learning (DRL)
multi-agent learning
期刊
IF:
3.6
论文数:
9.8W
被引数:
29.4W
机构
引用论文
The Binding of DYNLL2 to Myosin Va Requires Alternatively Spliced Exon B and Stabilizes a Portion of the Myosin's Coiled-Coil Domain
Biochemistry
IF0
Multi-Flow Transmission in Wireless Interference Networks: A Convergent Graph Learning Approach无线干扰网络中的多流传输: 一种收敛图学习方法
Selected Methods for the Oxidation of 1,1,1-Trichloro-2-alkanols. An Efficient Modification Using Chromic Acid
Synthesis
IF0

