arrow
Return

Wireless Resource Optimization for UAV Swarm Cooperative Sensing via Multiagent Multitask Deep Reinforcement Learning

delete2026-01-22
delete0
PRE
AI
M
Mengjie Li
A
Anming Dong
G
Guijuan Wang
X
Xiang Tian
S
Sufang Li
J
Jiguo Yu
DOI:10.1109/JIOT.2026.3657123delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Cooperative sensing through multiple uncrewed aerial vehicles (UAV) swarms is a promising solution for many critical scenarios such as disaster monitoring, environmental surveillance, and military operations. Optimizing wireless resources, including spectrum and power allocation, is essential in UAV swarm systems to maximize their utilization. In this work, we investigate wireless resource optimization for a latency-sensitive multiswarm coordination system based on deep reinforcement learning (DRL). A multiobjective joint subchannel and power allocation problem is formulated to minimize the expected age of information (AoI) and power consumption for each swarm, while maximizing the transmission probability of cooperative awareness messages (CAMs). The multiagent DRL (MADRL) framework is employed to simultaneously and effectively optimize multiple objectives, with each swarm leader (SL) acting as an independent agent capable of dynamically interacting with its environment to learn the optimal policy. Two distinct critic networks are designed to tackle the dual challenges of promoting cooperation among agents while also enhancing individual performance. A global critic is implemented to assess the system-wide expected reward, thereby encourage coordinated behaviors among agents. In parallel, local critics are tailored for each individual agent, focusing on evaluating agent-specific rewards. We also adopt a task-aware rewarding mechanism by decomposing the individual reward of each agent into multiple subreward functions based on the tasks each agent needs to accomplish, enabling the learning of task-specific value functions separately. The simulation results show that the proposed algorithm outperforms previous DRL frameworks in terms of AoI and CAM transmission.
Keywords:
Age of information (AoI)
multiagent deep reinforcement learning (MADRL)
resource management
task decomposition
uncrewed aerial vehicles (UAV) swarm cooperation

Journal

IEEE Internet of Things Journal cover
IEEE Internet of Things Journal
IF:
8.9
Papers:
1.4W
Citations:
7.8W

Organization

U
university of electronic science and technology of china
Scholars:
1.2W
Papers: 4.5K
Citations: 4
Q
qilu university of technology
Scholars:
2.0K
Papers: 605
Citations: 0
researcher View more organizations