返回
A State-Decomposition DDPG Algorithm for UAV Autonomous Navigation in 3-D Complex Environments
DOI:10.1109/JIOT.2023.3327753.png)
摘要
En 中文
Over the past decade, unmanned aerial vehicles (UAVs) have been widely applied in many areas, such as goods delivery, disaster monitoring, search and rescue etc. In most of these applications, autonomous navigation is one of the key techniques that enable UAV to perform various tasks. However, UAV autonomous navigation in complex environments presents significant challenges due to the difficulty in simultaneously observing, orientation, decision and action. In this work, an efficient state-decomposition deep deterministic policy gradient algorithm is proposed for UAV autonomous navigation (SDDPG-NAV) in 3-D complex environments. In SDDPG-NAV, a novel state-decomposition method that uses two subnetworks for the perception-related and target-related states separately is developed to establish more appropriate actor networks. We also designed some objective-oriented reward functions to solve the sparse reward problem, including approaching the target, and avoiding obstacles and step award functions. Moreover, some training strategies are introduced to maintain the balance between exploration and exploitation, and the network is well trained with numerous experiments. The proposed SDDPG-NAV algorithm is capable of adapting to surrounding environments with generalized training experiences and effectively improves UAV's navigation performance in 3-D complex environments. Comparing with the benchmark DDPG and TD3 algorithms, SDDPG-NAV exhibits better performance in terms of convergence rate, navigation performance, and generalization capability.
Keyword:
Autonomous aerial vehicles
Navigation
Autonomous robots
Three-dimensional displays
Training
Heuristic algorithms
Internet of Things
Autonomous navigation
decision making
deep reinforcement learning (DRL)
path planning
unmanned aerial vehicle (UAV) autonomy
期刊
IF:
8.9
论文数:
1.4W
被引数:
7.8W
机构
暂无机构信息
引用论文
A UAV Navigation Approach Based on Deep Reinforcement Learning in Large Cluttered 3D Environments基于深度强化学习的大型三维环境下无人机导航方法
Deep-Reinforcement-Learning-Based Autonomous UAV Navigation With Sparse Rewards具有稀疏奖励的基于深度强化学习的自主无人机导航
Memory-Based Deep Reinforcement Learning for Obstacle Avoidance in UAV With Limited Environment Knowledge基于记忆的深度强化学习在有限环境知识下的无人机避障
A DDPG-based Approach for Energy-aware UAV Navigation in Obstacle-constrained Environment障碍物约束环境下基于DDPG的能量感知无人机导航方法

