1
Return

Improving Energy Efficiency in Post-Disaster Networks: Multi-Agent Deep Reinforcement Learning for Enhancing Local Contributions

delete2026-02-18
delete0
PRE
AI
Y
Y. Y. Dang
迟学芬 cover
迟学芬 (Xuefen Chi)
H
Hangyu Yan
T
Tianyue Zhang
X
Xin Zhang
Z
Zhu Han
DOI:10.1109/tvt.2026.3666157delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
To address the energy efficiency (EE) optimization challenge in post-disaster space-air enhanced integrated access and backhaul networks (SAE-IABN), a multi-agent global trust region local policy optimization (MA-GTRLPO) based resource allocation (RA) strategy is proposed. This strategy adopts a hybrid hierarchical architecture, decoupling global RA into centralized authorization and decentralized allocation to manage the multi-layer structured SAE-IABN’s resources efficiently. MA-GTRLPO enhances local contributions through a sequential local policy update rule, which helps stabilize convergence and avoids joint policy degradation risk. The general global performance difference bound theorem provides theoretical convergence guarantees for MA-GTRLPO. This theorem also offers a coordination mechanism for the Kullback-Leibler (KL) regularization coefficient and learning rate, improving global rewards by 6.9% compared to a heuristic hyperparameter baseline. Next, the global reward with the dynamic feedback framework and triple normalization is proposed to optimize EE further. The feedback framework with quality of service (QoS) constraints is introduced to prevent over-allocation. Then, the triple normalization helps agents clarify the contribution of EE, QoS guarantee, and energy constraint to the global reward. Simulations demonstrate that MA-GTRLPO outperforms the baselines in all metrics. It shows good scalability, robustness, and timeliness under varying network densities and topologies, guaranteeing QoS for user equipment (UE) and ensuring the endurance of uncrewed aerial vehicles (UAVs) remains above the minimum endurance bound.
Keywords:
Resource allocation
quality of service
multi-agent deep reinforcement learning
post-disaster communication

Journal

IEEE Transactions on Vehicular Technology cover
IEEE Transactions on Vehicular Technology
IF:
7.1
Papers:
1.7W
Citations:
6.6W

Organization

U
University of Houston
Scholars:
921
Papers: 535
Citations: 1.7W
J
Jilin University
Scholars:
8.4W
Papers: 5.5W
Citations: 8.9K
Cited Papers

Cited Papers

Citing Papers

Citing Papers