arrow
返回

Multi-Agent Deep Reinforcement Learning for Dynamic Power Allocation in Wireless Networks

delete2019-10-01
delete413
delete
OA
AI
Y
Yasar Sinan Nasir *
D
Dongning Guo
DOI:10.1109/JSAC.2019.2933973delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
This work demonstrates the potential of deep reinforcement learning techniques for transmit power control in wireless networks. Existing techniques typically find near-optimal power allocations by solving a challenging optimization problem. Most of these algorithms are not scalable to large networks in real-world scenarios because of their computational complexity and instantaneous cross-cell channel state information (CSI) requirement. In this paper, a distributively executed dynamic power allocation scheme is developed based on model-free deep reinforcement learning. Each transmitter collects CSI and quality of service (QoS) information from several neighbors and adapts its own transmit power accordingly. The objective is to maximize a weighted sum-rate utility function, which can be particularized to achieve maximum sum-rate or proportionally fair scheduling. Both random variations and delays in the CSI are inherently addressed using deep Q-learning. For a typical network architecture, the proposed algorithm is shown to achieve near-optimal power allocation in real time based on delayed CSI measurements available to the agents. The proposed scheme is especially suitable for practical scenarios where the system model is inaccurate and CSI delay is non-negligible.
Keyword:
Deep Q-learning
radio resource management
interference mitigation
power control
Jakes fading model
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Journal on Selected Areas in Communications 封面图
IEEE Journal on Selected Areas in Communications
IF:
17.2
论文数:
6.4K
被引数:
3.1W

机构

N
Northwestern University
学者数:
6.2W
论文数: 5.3W
被引数: 3.9K
引用论文

引用论文

A free-floating bike repositioning problem with faulty bikes
err2019-01-01
err0
errOAAI
errMuhammad Usama; Yongjun Shen; Onaira Zahoor
err分享
err收藏
Evaluating more naturalistic outcome measures评估更自然的结果测量
err2015-12-01
err0
errOAAI
errRiley Bove; Charles C. White; Gavin Giovannoni; Bonnie Glanz; Victor Golubchikov; Johnny Hujol; Charles Jennings; Dawn Langdon; Michelle Lee; Anna Legedza; James Paskavitz; Sashank Prasad; John Richert; Allison Robbins; Susan Roberts; Howard Weiner; Ravi Ramachandran; Martyn Botfield; Philip L. De Jager
err分享
err收藏
Microbial modulation of behavior and stress responses in zebrafish larvae
err2016-09-01
err0
errOAAI
errDaniel J. Davis; Elizabeth C. Bryda; Catherine H. Gillespie; Aaron C. Ericsson
err分享
err收藏
Weighted Sum-Rate Maximization in Multi-Cell Networks via Coordinated Scheduling and Discrete Power Control
err2011-06-01
err126
PREAI
errZhang, Honghai; Venturino, Luca; Prasad, Narayan; Li, Peilong; Rangarajan, Sampath; Wang, Xiaodong
err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容