返回
Modular Reinforcement Learning for Autonomous UAV Flight Control
DOI:10.3390/drones7070418.png)
摘要
En 中文
Recently, research on unmanned aerial vehicles (UAVs) has increased significantly. UAVs do not require pilots for operation, and UAVs must possess autonomous flight capabilities to ensure that they can be controlled without a human pilot on the ground. Previous studies have mainly focused on rule-based methods, which require specialized personnel to create rules. Reinforcement learning has been applied to research on UAV autonomous flight; however, it does not include six-degree-of-freedom (6-DOF) environments and lacks realistic application, resulting in difficulties in performing complex tasks. This study proposes a method of efficient learning by connecting two different maneuvering methods using modular learning for autonomous UAV flights. The proposed method divides complex tasks into simpler tasks, learns them individually, and then connects them in order to achieve faster learning by transferring information from one module to another. Additionally, the curriculum learning concept was applied, and the difficulty level of individual tasks was gradually increased, which strengthened the learning stability. In conclusion, modular learning and curriculum learning methods were used to demonstrate that UAVs can effectively perform complex tasks in a realistic, 6-DOF environment.
Keyword:
UAV
autonomous flight control
reinforcement learning
modular learning
curriculum learning
JSBSim
期刊
D
IF:
4.8
论文数:
3.9K
被引数:
8.3K
机构
引用论文
No association of a tyrosine hydroxylase gene tetranucleotide repeat polymorphism in autism, Tourette syndrome, or ADHD自闭症,Tourette综合征或ADHD中酪氨酸羟化酶基因四核苷酸重复多态性无关联
Autonomous Control of Combat Unmanned Aerial Vehicles to Evade Surface-to-Air Missiles Using Deep Reinforcement Learning利用深度强化学习自主控制作战无人机规避地空导弹
IEEE ACCESS
IF3.6
Hierarchical Fully Convolutional Network for Joint Atrophy Localization and Alzheimer's Disease Diagnosis Using Structural MRI使用结构MRI进行关节萎缩定位和阿尔茨海默氏病诊断的分层完全卷积网络
How to use the diffusion model: Parameter recovery of three methods: EZ, fast-dm, and DMAT如何使用扩散模型: 三种方法的参数恢复: EZ,fast-dm和DMAT

