返回
Learning-Based Optimal Cooperative Formation Tracking Control for Multiple UAVs: A Feedforward-Feedback Design Framework
DOI:10.1109/TASE.2023.3322028.png)
摘要
En 中文
Notwithstanding the successful design of state-of-the-art cooperative control protocols to accomplish formation tracking for multiple unmanned aerial vehicles (UAVs), the assurance of performance optimality cannot be guaranteed in the face of complex disturbances affecting these multi-UAV systems. In order to surmount this challenge, this research endeavor aims to establish a feedforward-feedback learning-based optimal control methodology to facilitate cooperative UAV formation tracking in the presence of intricate disturbances. To be more precise, by leveraging backstepping-based feedback control, the problem of UAV formation tracking is transformed into an equivalent optimal regulation problem. Consequently, a learning-based feedforward control scheme is devised, wherein the cooperative policy iteration algorithm is formulated based on a two-player zero-sum game. The critic-only echo state network (ESN) is employed to approximate the optimal feedforward control policies, with the inclusion of an online adaptive tuning law and compensation terms to alleviate the persistence of excitation condition and eliminate the need for an initial admissible control. As a result, the closed-loop stability is guaranteed in terms of uniformly ultimately boundedness for tracking errors and ESN weights. Note to Practitioners-In real-world scenarios, the flight of multiple UAVs is invariably affected by intricate disturbances, resulting in compromised tracking precision. There is an urgent need to enhance resistance to disturbances and ensure optimal performance for cooperative formation tracking of multiple UAVs. Beyond the capabilities of model-based controllers, the integration of reinforcement learning has shown promise in achieving robust control actions. By introducing the cooperative policy iteration algorithm based on a two-player zero-sum game, the tracking performances of UAV formation can be further optimized. In order to facilitate the practical application of reinforcement learning in UAV systems, our proposed algorithm addresses the persistency of excitation condition by incorporating innovative compensation terms into the ESN tuning law. Furthermore, we resolve the requirement for initial admissible control by introducing a novel piecewise compensation term into the ESN tuning law, which is based on a newly proposed Lyapunov function.
Keyword:
Feedforward-feedback learning-based control
two-player zero-sum game
unmanned aerial vehicle formation tracking
期刊
IF:
6.4
论文数:
5.1K
被引数:
1.6W
机构
引用论文
An intelligent cooperative mission planning scheme of UAV swarm in uncertain dynamic environment不确定动态环境下无人机群智能协同任务规划方案
Adaptive Critic Learning and Experience Replay for Decentralized Event-Triggered Control of Nonlinear Interconnected Systems非线性互联系统分散事件触发控制的自适应批评家学习和经验重放
Time-Varying Formation Control for Unmanned Aerial Vehicles: Theories and Applications无人机时变编队控制: 理论与应用
Decentralized Event-Triggered Control for a Class of Nonlinear-Interconnected Systems Using Reinforcement Learning基于强化学习的一类非线性关联系统的分散事件触发控制
Memory-Based Deep Reinforcement Learning for Obstacle Avoidance in UAV With Limited Environment Knowledge基于记忆的深度强化学习在有限环境知识下的无人机避障
Reinforcement Learning-Based Optimal Tracking Control of an Unknown Unmanned Surface Vehicle基于强化学习的未知无人水面舰艇最优跟踪控制
Time-Varying Formation Tracking for UAV Swarm Systems With Switching Directed Topologies切换定向拓扑的无人机群系统时变编队跟踪

