arrow
Return

Dynamic threshold-enhanced diffusion PPO for multi-UAV collaborative optimization in wireless rechargeable sensor networks

delete2026-07-04
delete0
delete
OA
AI
Y
Yalin Nie
孙泽钰 cover
孙泽钰 (Zeyu Sun) *
Y
Yang Zhang
DOI:10.1038/s41598-026-60774-6delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
To address the challenging problem of collaborative optimization of communication delay and UAV load balancing in multi-Unmanned Aerial Vehicle (UAV)-assisted wireless rechargeable sensor networks, a dynamic threshold-enhanced diffusion proximal policy optimization algorithm (DTD-PPO) is proposed. Firstly, a multi-objective optimization model of multi-UAV-assisted WRSNs is constructed, and multi-dimensional constraints are incorporated to enhance the feasibility and practicality of the optimization solution. Secondly, a Markov Decision Process (MDP) framework is designed to balance the conflict between the dual objectives through dynamic weighting. To improve the exploration ability and training stability of the algorithm, the diffusion model is integrated into the PPO policy network, generating diversified actions through an adaptive noise-adding and denoising process. Additionally, a dynamic threshold strategy based on the normalized reward change rate is proposed to adjust the policy update magnitude in real-time. The effectiveness of our proposed algorithm is validated by using metrics of the data collection delay, UAV’s flight distance deviation and the energy efficiency. The simulation results verify the superiority and robustness of DTD-PPO algorithm compared to the other benchmark methods.

Journal

Scientific Reports cover
Scientific Reports
IF:
3.9
Papers:
27.1W
Citations:
83.5W

Organization

L
Luoyang Institute of Science and Technology
Scholars:
401
Papers: 164
Citations: 1.8K