返回
Deep Multi-User Reinforcement Learning for Distributed Dynamic Spectrum Access
DOI:10.1109/TWC.2018.2879433.png)
摘要
En 中文
We consider the problem of dynamic spectrum access for network utility maximization in multichannel wireless networks. The shared bandwidth is divided into K orthogonal channels. In the beginning of each time slot, each user selects a channel and transmits a packet with a certain transmission probability. After each time slot, each user that has transmitted a packet receives a local observation indicating whether its packet was successfully delivered or not (i.e., ACK signal). The objective is a multi-user strategy for accessing the spectrum that maximizes a certain network utility in a distributed manner without online coordination or message exchanges between users. Obtaining an optimal solution for the spectrum access problem is computationally expensive, in general, due to the large-state space and partial observability of the states. To tackle this problem, we develop a novel distributed dynamic spectrum access algorithm based on deep multi-user reinforcement leaning. Specifically, at each time slot, each user maps its current state to the spectrum access actions based on a trained deep-Q network used to maximize the objective function. Game theoretic analysis of the system dynamics is developed for establishing design principles for the implementation of the algorithm. The experimental results demonstrate the strong performance of the algorithm.
Keyword:
Wireless networks
dynamic spectrum access
medium access control (MAC) protocols
multi-agent learning
deep reinforcement learning
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
10.7
论文数:
1.3W
被引数:
5.3W
机构
引用论文
[AEPH2][GeSb2S6]·CH3OH: a thiogermanate–thioantimonate featuring an infinite ribbon-like structure with an unusual {GeSb3S11} unit and exhibiting the ability of photocatalytic degradation of organic dye
CrystEngComm
IF0
The Binding of DYNLL2 to Myosin Va Requires Alternatively Spliced Exon B and Stabilizes a Portion of the Myosin's Coiled-Coil Domain
Biochemistry
IF0

