arrow
返回

Learning-Based Optimal Cooperative Formation Tracking Control for Multiple UAVs: A Feedforward-Feedback Design Framework

delete2025-01-01
delete2
PRE
AI
B
Boyang Zhang
M
Maolong Lv *
S
Shaohua Cui
X
Xiangwei Bu
J
Ju H. Park
DOI:10.1109/TASE.2023.3322028delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Notwithstanding the successful design of state-of-the-art cooperative control protocols to accomplish formation tracking for multiple unmanned aerial vehicles (UAVs), the assurance of performance optimality cannot be guaranteed in the face of complex disturbances affecting these multi-UAV systems. In order to surmount this challenge, this research endeavor aims to establish a feedforward-feedback learning-based optimal control methodology to facilitate cooperative UAV formation tracking in the presence of intricate disturbances. To be more precise, by leveraging backstepping-based feedback control, the problem of UAV formation tracking is transformed into an equivalent optimal regulation problem. Consequently, a learning-based feedforward control scheme is devised, wherein the cooperative policy iteration algorithm is formulated based on a two-player zero-sum game. The critic-only echo state network (ESN) is employed to approximate the optimal feedforward control policies, with the inclusion of an online adaptive tuning law and compensation terms to alleviate the persistence of excitation condition and eliminate the need for an initial admissible control. As a result, the closed-loop stability is guaranteed in terms of uniformly ultimately boundedness for tracking errors and ESN weights. Note to Practitioners-In real-world scenarios, the flight of multiple UAVs is invariably affected by intricate disturbances, resulting in compromised tracking precision. There is an urgent need to enhance resistance to disturbances and ensure optimal performance for cooperative formation tracking of multiple UAVs. Beyond the capabilities of model-based controllers, the integration of reinforcement learning has shown promise in achieving robust control actions. By introducing the cooperative policy iteration algorithm based on a two-player zero-sum game, the tracking performances of UAV formation can be further optimized. In order to facilitate the practical application of reinforcement learning in UAV systems, our proposed algorithm addresses the persistency of excitation condition by incorporating innovative compensation terms into the ESN tuning law. Furthermore, we resolve the requirement for initial admissible control by introducing a novel piecewise compensation term into the ESN tuning law, which is based on a newly proposed Lyapunov function.
Keyword:
Feedforward-feedback learning-based control
two-player zero-sum game
unmanned aerial vehicle formation tracking

期刊

IEEE Transactions on Automation Science and Engineering 封面图
IEEE Transactions on Automation Science and Engineering
IF:
6.4
论文数:
5.1K
被引数:
1.6W

机构

B
Beihang University
学者数:
5.2W
论文数: 4.1W
被引数: 37
A
Air Force Engineering University
学者数:
4.8K
论文数: 3.0K
被引数: 1.9K
Y
Yeungnam University
学者数:
1.0W
论文数: 1.3W
被引数: 1.4W
学者 查看更多机构
引用论文

引用论文

err分享
err收藏
学者 查看更多内容