返回
Deep Reinforcement Learning and Markov Decision Problem for Task Offloading in Mobile Edge Computing
DOI:10.1007/s10723-023-09708-4.png)
摘要
En 中文
Mobile Edge Computing (MEC) offers cloud-like capabilities to mobile users, making it an up-and-coming method for advancing the Internet of Things (IoT). However, current approaches are limited by various factors such as network latency, bandwidth, energy consumption, task characteristics, and edge server overload. To address these limitations, this research propose a novel approach that integrates Deep Reinforcement Learning (DRL) with Deep Deterministic Policy Gradient (DDPG) and Markov Decision Problem for task offloading in MEC. Among DRL algorithms, the ITODDPG algorithm based on the DDPG algorithm and MDP is a popular choice for task offloading in MEC. Firstly, the ITODDPG algorithm formulates the task offloading problem in MEC as an MDP, which enables the agent to learn a policy that maximizes the expected cumulative reward. Secondly, ITODDPG employs a deep neural network to approximate the Q-function, which maps the state-action pairs to their expected cumulative rewards. Finally, the experimental results demonstrate that the ITODDPG algorithm outperforms the baseline algorithms regarding average compensation and convergence speed. In addition to its superior performance, our proposed approach can learn complex non-linear policies using DNN and an information-theoretic objective function to improve the performance of task offloading in MEC. Compared to traditional methods, our approach delivers improved performance, making it highly effective for developing IoT environments. Experimental trials were carried out, and the results indicate that the suggested approach can enhance performance compared to the other three baseline methods. It is highly scalable, capable of handling large and complex environments, and suitable for deployment in real-world scenarios, ensuring its widespread applicability to a diverse range of task offloading and MEC applications.
Keyword:
Deep Reinforcement Learning
Deep Deterministic Policy Gradient
Mobile Edge Computing
Task offloading
Markov decision problem
期刊
IF:
2.9
论文数:
763
被引数:
1.2K
机构
引用论文
Security defense decision method based on potential differential game for complex networks基于势差博弈的复杂网络安全防御决策方法
COMPUTERS & SECURITY
IF5.4
System Dynamics Approach for Evaluating the Interconnection Performance of Cross-Border Transport Infrastructure跨境运输基础设施互联互通性能评价的系统动力学方法
Adaptive Dynamic Surface Control With Disturbance Observers for Battery/Supercapacitor-Based Hybrid Energy Sources in Electric Vehicles电动汽车中基于电池/超级电容器的混合能源的带有干扰观测器的自适应动态表面控制

