arrow
返回

Multi-Agent Reinforcement Learning-Based Resource Allocation for UAV Networks

delete2020-02-01
delete352
delete
OA
AI
J
Jingjing Cui *
Y
Yuanwei Liu
A
Arumugam Nallanathan
DOI:10.1109/TWC.2019.2935201delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Unmanned aerial vehicles (UAVs) are capable of serving as aerial base stations (BSs) for providing both cost-effective and on-demand wireless communications. This article investigates dynamic resource allocation of multiple UAVs enabled communication networks with the goal of maximizing long-term rewards. More particularly, each UAV communicates with a ground user by automatically selecting its communicating user, power level and subchannel without any information exchange among UAVs. To model the dynamics and uncertainty in environments, we formulate the long-term resource allocation problem as a stochastic game for maximizing the expected rewards, where each UAV becomes a learning agent and each resource allocation solution corresponds to an action taken by the UAVs. Afterwards, we develop a multi-agent reinforcement learning (MARL) framework that each agent discovers its best strategy according to its local observations using learning. More specifically, we propose an agent-independent method, for which all agents conduct a decision algorithm independently but share a common structure based on Q-learning. Finally, simulation results reveal that: 1) appropriate parameters for exploitation and exploration are capable of enhancing the performance of the proposed MARL based resource allocation algorithm; 2) the proposed MARL algorithm provides acceptable performance compared to the case with complete information exchanges among UAVs. By doing so, it strikes a good tradeoff between performance gains and information exchange overheads.
Keyword:
Resource management
Trajectory
Wireless communication
Communication networks
Dynamic scheduling
Stochastic processes
Reinforcement learning
Dynamic resource allocation
multi-agent reinforcement learning (MARL)
stochastic games
UAV communications
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Wireless Communications 封面图
IEEE Transactions on Wireless Communications
IF:
10.7
论文数:
1.3W
被引数:
5.3W

机构

U
university of southampton
学者数:
3.3W
论文数: 3.2W
被引数: 52
U
university of london
学者数:
21.5W
论文数: 19.7W
被引数: 305
引用论文

引用论文

err分享
err收藏
err分享
err收藏
err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容