返回
Online estimation of objective function for continuous-time deterministic systems
DOI:10.1016/j.neunet.2024.106116.png)
摘要
En 中文
We developed two online data -driven methods for estimating an objective function in continuous -time linear and nonlinear deterministic systems. The primary focus addressed the challenge posed by unknown input dynamics (control mapping function) in the expert system, a critical element for an online solution of the problem. Our methods leverage both the learner's and expert's data for effective problem -solving. The first approach, which is model -free, estimates the expert's policy and integrates it into the learner agent to approximate the objective function associated with the optimal policy. The second approach estimates the input dynamics from the learner's data and combines it with the expert's input -state observations to tackle the objective function estimation problem. Compared to other methods for deterministic systems that rely on both the learner's and expert's data, our approaches offer reduced complexity by eliminating the need to estimate an optimal policy after each objective function update. We conduct a convergence analysis of the estimation techniques using Lyapunov-based methods. Numerical experiments validate the effectiveness of our developed methods.
Keyword:
Objective function estimation
Deterministic systems
Data-driven solution
Continuous-time systems
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.3
论文数:
7.8K
被引数:
3.0W
机构
暂无机构信息
引用论文
Inverse reinforcement learning for multi-player noncooperative apprentice games多玩家非合作学徒游戏的逆强化学习
AUTOMATICA
IF5.9
Adaptive Optimal Control of Unknown Constrained-Input Systems Using Policy Iteration and Neural Networks基于策略迭代和神经网络的未知约束输入系统的自适应最优控制
A novel actor-critic-identifier architecture for approximate optimal control of uncertain nonlinear systems一种用于不确定非线性系统的近似最优控制的新型参与者-批评者-标识符体系结构
AUTOMATICA
IF5.9
From inverse optimal control to inverse reinforcement learning: A historical review从逆最优控制到逆强化学习: 历史回顾
Efficient model-based reinforcement learning for approximate online optimal control基于模型的有效强化学习用于近似在线最优控制
AUTOMATICA
IF5.9

