1
Return

Deep Reinforcement Learning Enabled Persistent Surveillance With Energy-Aware UAV-UGV Systems for Disaster Management Applications

delete2026-05-26
delete0
delete
OA
AI
M
Md Safwan Mondal
R
R. Subramanian
J
James Humann
J
James M. Dotterweich
P
Pranav A. Bhounsule
DOI:10.1109/tiv.2026.3696943delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Integrating Unmanned Aerial Vehicles (UAVs) with Unmanned Ground Vehicles (UGVs) provides an effective solution for persistent surveillance in disaster management. UAVs excel at covering large areas rapidly, but their range is limited by battery capacity. UGVs, though slower, can carry larger batteries for extended missions. By using UGVs as mobile recharging stations, UAVs can extend mission duration through periodic refueling, leveraging the complementary strengths of both systems. To optimize this energy-aware UAV-UGV cooperative routing problem, we propose a planning framework that determines optimal routes and recharging points between a UAV and a UGV. Our solution employs a deep reinforcement learning (DRL) framework built on an encoder-decoder transformer architecture with multi-head attention mechanisms. This architecture enables the model to sequentially select actions for visiting mission points and coordinating recharging rendezvous between the UAV and UGV. The DRL model is trained to minimize the <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">age periods</i> (the time gap between consecutive visits) of mission points, ensuring effective surveillance. We evaluate the framework across various problem sizes and distributions, comparing its performance against heuristic methods and an existing learning-based model. Results show that our approach consistently outperforms these baselines in both solution quality and runtime. Additionally, we demonstrate the DRL policy’s applicability in a real-world disaster scenario as a case study and explore its potential for online mission planning to handle dynamic changes. Adapting the DRL policy for priority-driven surveillance highlights the model’s generalizability for real-time disaster response.
Keywords:
Disaster management
deep reinforcement learning
UAV
UGV
surveillance
cooperative routing
multi-agent systems
combinatorial optimization

Journal

I
IEEE Transactions on Intelligent Vehicles
IF:
14.3
Papers:
1.2K
Citations:
1.2W

Organization

U
university of illinois chicago
Scholars:
1.8K
Papers: 886
Citations: 0
A
army research laboratory
Scholars:
86
Papers: 50
Citations: 0
Cited Papers

Cited Papers

Citing Papers

Citing Papers