arrow
返回

Multi-Agent Deep Reinforcement Learning-Based Interdependent Computing for Mobile Edge Computing-Assisted Robot Teams

delete2023-05-01
delete14
PRE
AI
Q
Qimei Cui *
赵奚誉 封面图
赵奚誉 (Xiyu Zhao)
W
Wei Ni
Z
Zheng Hu *
X
Xiaofeng Tao
张
张平 (Ping Zhang)
DOI:10.1109/TVT.2022.3232806delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
A group of robots can be assigned with different roles to collaboratively conduct interdependent tasks. The robots form a multi-robot system (MRS), where one robot's decision or action relies on the others'. This paper addresses the sequential decision problem of user association and resource allocation in a mobile edge computing (MEC)-enabled, wirelessly-connected MRS to maximize the time-averaged completion rate of interdependent computing tasks. The problem is challenging due to the partial observability of the network environment, and the delicate delay requirements of interdependent computing tasks. A new decentralized partially observable Markov decision process (Dec-POMDP) problem is reformulated, where edge servers act as intelligent agents and can make decentralized decisions about user association and resource management with their local information of the network state. By leveraging the multi-agent deep deterministic policy gradient (MADDPG) theory, a new cooperative multi-agent deep reinforcement learning (MADRL) model is developed to enable interdependent computing. Simulations show the merits of our approach in terms of task completion rate compared to existing techniques.
Keyword:
Robots
Task analysis
Servers
Resource management
Delays
Robot kinematics
Processor scheduling
Robot team
multi-robot system
mobile edge computing
interdependent computing
user association
resource allocation
multi-agent deep reinforcement learning

期刊

IEEE Transactions on Vehicular Technology 封面图
IEEE Transactions on Vehicular Technology
IF:
7.1
论文数:
1.8W
被引数:
6.6W

机构

B
beijing university of posts & telecommunications
学者数:
1.4W
论文数: 1.2W
被引数: 9
C
引用论文

引用论文

Electrohydrodynamic stability of a fluid layer. II. Effect of a normal electric field
err1986-07-01
err0
PREAI
errAbou El Magd A. Mohamed; El Sayed F. El Shehawey; Yusry O. El Dib
err分享
err收藏
MDP-Based Task Offloading for Vehicular Edge Computing Under Certain and Uncertain Transition Probabilities
err2020-03-01
err82
PREAI
errZhang, Xuefei; Zhang, Jian; Liu, Zhitong; Cui, Qimei; Tao, Xiaofeng; Wang, Shuo
err分享
err收藏
err分享
err收藏
IL-1 regulates the Cyp7a1 gene and serum total cholesterol level at steady state in mice
err2009-02-01
err0
PREAI
errMisaki Kojima; Takashi Ashino; Takemi Yoshida; Yoichiro Iwakura; Masashi Sekimoto; Masakuni Degawa
err分享
err收藏
Stochastic Online Learning for Mobile Edge Computing: Learning from Changes
err2019-03-01
err85
PREAI
errCui, Qimei; Gong, Zhenzhen; Ni, Wei; Hou, Yanzhao; Chen, Xiang; Tao, Xiaofeng; Zhang, Ping
err分享
err收藏
err分享
err收藏
Online Anticipatory Proactive Network Association in Mobile Edge Computing for IoT
err2020-07-01
err32
PREAI
errCui, Qimei; Zhang, Jian; Zhang, Xuefei; Chen, Kwang-Cheng; Tao, Xiaofeng; Zhang, Ping
err分享
err收藏
Preference, context and communities
err2013-09-08
err0
PREAI
errYe Xu; Mu Lin; Hong Lu; Giuseppe Cardone; Nicholas Lane; Zhenyu Chen; Andrew Campbell; Tanzeem Choudhury
err分享
err收藏
学者 查看更多内容