返回
Approximate Dynamic Programming for Nonlinear-Constrained Optimizations
DOI:10.1109/TCYB.2019.2926248.png)
摘要
En 中文
In this paper, we study the constrained optimization problem of a class of uncertain nonlinear interconnected systems. First, we prove that the solution of the constrained optimization problem can be obtained through solving an array of optimal control problems of constrained auxiliary subsystems. Then, under the framework of approximate dynamic programming, we present a simultaneous policy iteration (SPI) algorithm to solve the Hamilton-Jacobi-Bellman equations corresponding to the constrained auxiliary subsystems. By building an equivalence relationship, we demonstrate the convergence of the SPI algorithm. Meanwhile, we implement the SPI algorithm via an actor-critic structure, where actor networks are used to approximate optimal control policies and critic networks are applied to estimate optimal value functions. By using the least squares method and the Monte Carlo integration technique together, we are able to determine the weight vectors of actor and critic networks. Finally, we validate the developed control method through the simulation of a nonlinear interconnected plant.
Keyword:
Optimal control
Interconnected systems
Optimization
Nonlinear systems
Dynamic programming
Approximation algorithms
Approximate dynamic programming
constrained optimization
neural networks
nonlinear interconnected system
optimal control
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
10.5
论文数:
1.1W
被引数:
5.0W
机构
引用论文
Ac-Electrogravimetry Study of Electroactive Thin Films. II. Application to Polypyrrole电活性薄膜的交流电重分析研究。二。聚吡咯的应用
Discrete-Time Local Value Iteration Adaptive Dynamic Programming: Convergence Analysis离散局部值迭代自适应动态规划: 收敛性分析
Integral reinforcement learning and experience replay for adaptive optimal control of partially-unknown constrained-input continuous-time systems
AUTOMATICA
IF5.9
Reinforcement-Learning-Based Robust Controller Design for Continuous-Time Uncertain Nonlinear Systems Subject to Input Constraints基于强化学习的具有输入约束的连续时间不确定非线性系统的鲁棒控制器设计
Near Optimal Event-Triggered Control of Nonlinear Discrete-Time Systems Using Neurodynamic Programming使用神经动力学规划的非线性离散时间系统的近似最优事件触发控制

