arrow
返回

Deep Reinforcement Learning for Delay-Oriented IoT Task Scheduling in SAGIN

delete2021-02-01
delete197
PRE
AI
C
Conghao Zhou
W
Wen Wu
H
Hongli He
P
Peng Yang
F
Feng Lyu
N
Nan Cheng *
X
Xuemin Shen
DOI:10.1109/TWC.2020.3029143delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this article, we investigate a computing task scheduling problem in space-air-ground integrated network (SAGIN) for delay-oriented Internet of Things (IoT) services. In the considered scenario, an unmanned aerial vehicle (UAV) collects computing tasks from IoT devices and then makes online offloading decisions, in which the tasks can be processed at the UAV or offloaded to the nearby base station or the remote satellite. Our objective is to design a task scheduling policy that minimizes offloading and computing delay of all tasks given the UAV energy capacity constraint. To this end, we first formulate the online scheduling problem as an energy-constrained Markov decision process (MDP). Then, considering the task arrival dynamics, we develop a novel deep risk-sensitive reinforcement learning algorithm. Specifically, the algorithm evaluates the risk, which measures the energy consumption that exceeds the constraint, for each state and searches the optimal parameter weighing the minimization of delay and risk while learning the optimal policy. Extensive simulation results demonstrate that the proposed algorithm can reduce the task processing delay by up to 30% compared to probabilistic configuration methods while satisfying the UAV energy capacity constraint.
Keyword:
Task analysis
Internet of Things
Processor scheduling
Delays
Unmanned aerial vehicles
Heuristic algorithms
US Department of Transportation
Space-air-ground integrated network
IoT
edge computing
reinforcement learning
constrained MDP
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Wireless Communications 封面图
IEEE Transactions on Wireless Communications
IF:
10.7
论文数:
1.3W
被引数:
5.3W

机构

C
Central South University
学者数:
10.0W
论文数: 7.2W
被引数: 10.9W
X
Xidian University
学者数:
2.4W
论文数: 1.9W
被引数: 9.7K
U
University of Waterloo
学者数:
2.2W
论文数: 2.3W
被引数: 3.3W
Z
zhejiang university
学者数:
17.7W
论文数: 12.1W
被引数: 152
学者 查看更多机构
引用论文

引用论文

Convergence of Edge Computing and Deep Learning: A Comprehensive Survey边缘计算和深度学习的融合: 综合综述
err2020-01-01
err812
errOAAI
errWang, Xiaofei; Han, Yiwen; Leung, Victor C. M.; Niyato, Dusit; Yan, Xueqiang; Chen, Xu
err分享
err收藏
Mobile Edge Computing: A Survey移动边缘计算: 一项调查
err2018-02-01
err2.0K
errOAAI
errAbbas, Nasir; Zhang, Yan; Taherkordi, Amir; Skeie, Tor
err分享
err收藏
Air-Ground Integrated Mobile Edge Networks: Architecture, Challenges, and Opportunities
err2018-08-01
err255
errOAAI
errCheng, Nan; Xu, Wenchao; Shi, Weisen; Zhou, Yi; Lu, Ning; Zhou, Haibo; Shen, Xuemin (Sherman)
err分享
err收藏
err分享
err收藏
err分享
err收藏
Reinforcement Learning-Based Downlink Interference Control for Ultra-Dense Small Cells
err2020-01-01
err67
PREAI
errXiao, Liang; Zhang, Hailu; Xiao, Yilin; Wan, Xiaoyue; Liu, Sicong; Wang, Li-Chun; Poor, H. Vincent
err分享
err收藏
Dynamic Spectrum Access in Multi-Channel Cognitive Radio Networks
err2014-11-01
err95
PREAI
errZhang, Ning; Liang, Hao; Cheng, Nan; Tang, Yujie; Mark, Jon W.; Shen, Xuemin (Sherman)
err分享
err收藏
学者 查看更多内容