1
Return

Joint trajectory and spectrum allocation optimization for UAV in complex electromagnetic environments: A multi-branch DRL approach

delete2026-03-01
delete0
PRE
AI
Y
Yihui Ye
Y
Yu Zhang *
W
Wan, Boyu
W
Wang, Wei
Y
Yangyi Zhang
S
Song, Xieda
DOI:10.1016/j.phycom.2026.103076delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In complex electromagnetic environments, Unmanned Aerial Vehicles (UAVs) face the dual challenges of avoiding Non-Flying Zones (NFZs) and counteracting dynamic hybrid jamming during mission execution, which necessitates the joint optimization of trajectory and spectrum allocation. To solve the problem, an Adaptive Sampling multi-Branch Channel-Enhanced Double Deep Q-Network (ASBCE-DDQN) algorithm is proposed. The algorithm integrates three key innovations: (1) the adaptive experience replay mechanism that prioritizes crucial communication experiences through a dynamic sampling strategy; (2) the multi-branch network for communication optimization that processes heterogeneous states and designs independent output branches for action dimensions; and (3) the enhanced channel selection module that adaptively adjusts channel Q-values via dynamic weighting. Furthermore, a composite reward function is designed to provide precise guidance for the UAV in balancing trajectory and spectrum allocation. Simulation results demonstrate that the proposed algorithm achieves a mission success rate of 91.93%, a communication success rate of 97.22% and an average flight steps reduction to 109.74 compared with baselines, confirming its effectiveness in joint trajectory and spectrum allocation optimization.
Keywords:
UAV
DRL
Joint optimization
Trajectory
Spectrum allocation

Journal

Physical Communication cover
Physical Communication
IF:
2.2
Papers:
279
Citations:
2.6K

Organization

N
national university of defense technology - china
Scholars:
1.8W
Papers: 1.4W
Citations: 9
Cited Papers

Cited Papers

Citing Papers

Citing Papers