返回
Approximate-optimal control algorithm for constrained zero-sum differential games through event-triggering mechanism
DOI:10.1007/s11071-018-4713-0.png)
摘要
En 中文
This paper investigates the optimization problem of two-player zero-sum differential game with control constraints in the framework of event triggering. Relying on reinforcement learning, an adaptive dynamic programming algorithm is developed to approximate the optimal solution of zero-sum game, i.e., the saddle-point equilibrium. A single-network structure is adopted, wherein a critic neural network (NN) evaluates the action. First, the constrained Hamilton-Jacobi-Isaacs equation is mathematically derived in the presence of control constraints; the event-triggering mechanism is then incorporated to reduce calculations and actions. Then, based on the gradient-descent technique, a novel weight updating law is designed for the critic NN, which ensures the solution can converge to the optimal value online. Moreover, the stability of closed-loop system is guaranteed and the unfavorable Zeno behavior is excluded by calculating the theoretical minimum triggering interval. Finally, two numerical examples are provided to verify the reliability and effectiveness of proposed algorithm.
Keyword:
Zero-sum differential game
Control constraints
Event triggering
Adaptive dynamic programming (ADP)
Neural network (NN)
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6
论文数:
1.4W
被引数:
4.1W
机构
引用论文
Data-driven adaptive dynamic programming for continuous-time fully cooperative games with partially constrained inputs
NEUROCOMPUTING
IF6.5
A three-network architecture for on-line learning and optimization based on adaptive dynamic programming
NEUROCOMPUTING
IF6.5
An adaptive critic neural network for motion control of a wheeled mobile robot用于轮式移动机器人运动控制的自适应评论家神经网络
Optimal Control of Multilayer Discrete Event Systems With Real-Time Constraint Guarantees具有实时约束保证的多层离散事件系统的最优控制

