arrow
返回

Multi-UAV roundup strategy method based on deep reinforcement learning CEL-MADDPG algorithm

delete2024-07-01
delete3
PRE
AI
B
Bo Li *
C
Chao Song
Z
Zhipeng Yang
K
Kaifang Wan
Q
Qingfu Zhang
DOI:10.1016/j.eswa.2023.123018delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Due to the complexity of the multi-UAV rounding up maneuvering target task in continuous and complex environments, it is difficult for the UAVs to quickly and accurately capture maneuvering targets. Therefore, this paper proposes CEL-MADDPG algorithm based on Curriculum Experience Learning. It improves the efficiency of multi-UAVs rounding up maneuvering target, and has certain generalization. which is better applied to the multiUAV roundup task in complex dynamic environments. The main contributions are the following two: By introducing the Curriculum Experience Learning, the multi-UAV rounding up task is divided into target tracking, encircling transition, and shrinking capture to learn, and designed corresponding reward function according to the task characteristics of each subtask. Which improves the learning efficiency of the model. Additionally, the CEL-MADDPG adopts the Preferential Experience Replay strategy to select experiences that are conducive to accelerating network convergence, and the experience most similar to the current state is further selected as a learning sample by using Relative Experience Learning (REL). This improves the sampling efficiency of samples and the training and optimization efficiency of the model. Simulation experiments show that the CEL-MADDPG algorithm can effectively improve the training efficiency of the model and has higher task completion efficiency.
Keyword:
Multi-UAV
CEL-MADDPG
Roundup
Deep-reinforcement

期刊

Expert Systems with Applications 封面图
Expert Systems with Applications
IF:
7.5
论文数:
2.9W
被引数:
10.2W

机构

C
City University of Hong Kong
学者数:
2.3W
论文数: 3.0W
被引数: 6.1W
N
Northwestern Polytechnical University
学者数:
4.6W
论文数: 3.7W
被引数: 5.3W
引用论文

引用论文

Application of remote technologies in education
err2019-04-08
err0
PREAI
errOlesya A. Stroeva; Yuliia Zviagintceva; Elena Tokmakova; Elena Petrukhina; Oksana Polyakova
err分享
err收藏
Multi-stage image denoising with the wavelet transform
err2023-02-01
err147
errOAAI
errTian, Chunwei; Zheng, Menghua; Zuo, Wangmeng; Zhang, Bob; Zhang, Yanning; Zhang, David
err分享
err收藏
On the synthesis of acrylamide oligomers
err2003-03-12
err0
PREAI
errEdgar Bortel; Andrzej Kochanowski; Antoni Kowalski
err分享
err收藏
Information consensus in multivehicle cooperative control
err2007-04-01
err2.8K
PREAI
errRen, Wei; Beard, Randal W.; Atkins, Ella M.
err分享
err收藏
Improving Autonomous Behavior Strategy Learning in an Unmanned Swarm System Through Knowledge Enhancement
err2022-06-01
err14
PREAI
errZhang, Tingting; Chai, Lai; Wang, Shenshen; Jin, Junyu; Liu, Xiaofan; Song, Aiguo; Lan, Yushi
err分享
err收藏
学者 查看更多内容