arrow
Return

Multi-UAV Cooperative Path Planning Method Based on an Improved MADDPG Algorithm

delete2026-04-17
delete0
delete
OA
AI
F
Feiqiao Zhang
王倩 cover
王倩 (Qian Wang)
X
Xin Ma *
DOI:10.3390/electronics15081632delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
To address cooperative path planning for multiple UAVs in complex environments, this paper proposes an improved multi-agent deep deterministic policy gradient algorithm, named Prioritized Experience Multi-Agent Deep Deterministic Policy Gradient (PE-MADDPG). An urban low-altitude inspection environment is first constructed within a reinforcement-learning framework, in which dynamic constraints, safety-separation requirements, and formation-cooperation objectives are incorporated into a partially observable Markov decision process. To improve training effectiveness, prioritized experience replay is introduced to increase the utilization of informative samples, an adaptive exploration-noise strategy is designed to regulate exploration intensity, and a multi-head attention mechanism is embedded in the Critic network to enhance the representation of inter-agent interactions. Simulation results in a three-dimensional urban inspection scenario show that PE-MADDPG outperforms the selected benchmark methods in task completion rate, formation maintenance, flight efficiency, and energy consumption. These results provide an effective solution for urban low-altitude inspection tasks.
Keywords:
multi-UAV
cooperative path planning
multi-agent reinforcement learning
prioritized experience replay
MADDPG
urban low-altitude inspection

Journal

Electronics cover
Electronics
IF:
2.6
Papers:
9.3K
Citations:
4.7W

Organization

C
civil aviation flight university of china
Scholars:
1.6K
Papers: 877
Citations: 0