arrow
返回

Deep-Deterministic Policy Gradient Based Multi-Resource Allocation in Edge-Cloud System: A Distributed Approach

delete2023-01-01
delete10
delete
OA
AI
A
Arslan Qadeer *
M
Myung Lee
DOI:10.1109/ACCESS.2023.3249153delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Edge Cloud (EC) empowers the beyond 5G (B5G) wireless networks to cope with large-scale and real-time traffics of Internet-of-Things (IoT) by minimizing the latency and providing compute power at the edge of the network. Due to a limited amount of resources at the EC compared to the back-end cloud (BC), intelligent resource management techniques become imperative. This paper studies the problem of multi-resource allocation (MRA) in terms of compute and wireless resources in an integrated EC and BC environment. Machine learning-based approaches are emerging to solve such optimization problems. However, it is challenging to adopt traditional discrete action space-based methods due to their high dimensionality issue. To this end, we propose a deep-deterministic policy gradient (DDPG) based temporal feature learning attentional network (TFLAN) model to address the MRA problem. TFLAN combines convolution, gated recurrent unit and attention layers together to mine local and long term temporal information from the task sequences for excellent function approximation. A novel heuristic-based priority experience replay (hPER) method is formulated to accelerate the convergence speed. Further, a pruning principle helps the TFLAN agent to significantly reduce the computational complexity and balance the load among base stations and servers to minimize the rejection-rate. Lastly, data parallelism technique is adopted for distributed training to meet the needs of a high-volume of IoT traffic in the EC environment. Experimental results demonstrate that the distributed training approach suites well to the problem scale and can magnify the speed of the learning process. We validate the proposed framework by comparing with five state-of-the-art RL agents. Our proposed agent converges fast and achieves up to 28% and 72% reduction in operational cost and rejection-rate, and achieves up to 32% gain in the quality of experience on average, compared to the most advanced DDPG agent.
Keyword:
Resource management
Internet of Things
Edge computing
Costs
Computational modeling
5G mobile communication
Cloud computing
Heuristic algorithms
Wireless networks
Edge cloud computing
wireless networks
deep deterministic policy gradient
resource allocation
smart city
IoT
beyond 5G
distributed training
heuristic priority experience replay

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

C
city university of new york (cuny) system
学者数:
1.6W
论文数: 1.5W
被引数: 26
引用论文

引用论文

err分享
err收藏
Occupational radiation doses in interventional radiology: simulations
err2008-02-18
err0
PREAI
errT. Siiskonen; M. Tapiovaara; A. Kosunen; M. Lehtinen; E. Vartiainen
err分享
err收藏
err分享
err收藏
Convergence of Edge Computing and Deep Learning: A Comprehensive Survey边缘计算和深度学习的融合: 综合综述
err2020-01-01
err812
errOAAI
errWang, Xiaofei; Han, Yiwen; Leung, Victor C. M.; Niyato, Dusit; Yan, Xueqiang; Chen, Xu
err分享
err收藏
Increasing Momentum-Like Factors: A Method for Reducing Training Errors on Multiple GPUs
err2022-02-01
err3
errOAAI
errTang, Yu; Kan, Zhigang; Yin, Lujia; Lai, Zhiquan; Zhang, Zhaoning; Qiao, Linbo; Li, Dongsheng
err分享
err收藏
学者 查看更多内容