arrow
Return

Online estimation of objective function for continuous-time deterministic systems

delete2024-04-01
delete2
delete
OA
AI
H
Hamed Jabbari Asl *
E
Eiji Uchibe
DOI:10.1016/j.neunet.2024.106116delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
We developed two online data -driven methods for estimating an objective function in continuous -time linear and nonlinear deterministic systems. The primary focus addressed the challenge posed by unknown input dynamics (control mapping function) in the expert system, a critical element for an online solution of the problem. Our methods leverage both the learner's and expert's data for effective problem -solving. The first approach, which is model -free, estimates the expert's policy and integrates it into the learner agent to approximate the objective function associated with the optimal policy. The second approach estimates the input dynamics from the learner's data and combines it with the expert's input -state observations to tackle the objective function estimation problem. Compared to other methods for deterministic systems that rely on both the learner's and expert's data, our approaches offer reduced complexity by eliminating the need to estimate an optimal policy after each objective function update. We conduct a convergence analysis of the estimation techniques using Lyapunov-based methods. Numerical experiments validate the effectiveness of our developed methods.
Keywords:
Objective function estimation
Deterministic systems
Data-driven solution
Continuous-time systems
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neural Networks cover
Neural Networks
IF:
6.3
Papers:
8.2K
Citations:
3.0W

Organization

No organization information available
Cited Papers

Cited Papers

Inverse reinforcement learning for multi-player noncooperative apprentice games
err2022-11-01
err30
errOAAI
errLian, Bosen; Xue, Wenqian; Lewis, Frank L.; Chai, Tianyou
errShare
errSave
A novel actor-critic-identifier architecture for approximate optimal control of uncertain nonlinear systems
err2013-01-01
err481
PREAI
errBhasin, S.; Kamalapurkar, R.; Johnson, M.; Vamvoudakis, K. G.; Lewis, F. L.; Dixon, W. E.
errShare
errSave
From inverse optimal control to inverse reinforcement learning: A historical review
err2020-01-01
err68
PREAI
errAb Azar, Nematollah; Shahmansoorian, Aref; Davoudi, Mohsen
errShare
errSave
Detecting Changes in Water Quality Data
err2008-01-01
err0
PREAI
errSean A. McKenna; Mark Wilson; Katherine A. Klise
errShare
errSave
Efficient model-based reinforcement learning for approximate online optimal control
err2016-12-01
err66
errOAAI
errKamalapurkar, Rushikesh; Rosenfeld, Joel A.; Dixon, Warren E.
errShare
errSave
AUTOMATED CALIBRATION STAMP TECHNOLOGY FOR IMPROVED IN‐SEASON NITROGEN FERTILIZATION
err2005-01-01
err0
PREAI
errW. R. Raun; J. B. Solie; M. L. Stone; D. L. Zavodny; K. L. Martin; K. W. Freeman
errShare
errSave
Genomic and transcriptomic correlates of Richter transformation in chronic lymphocytic leukemia
err2021-05-20
err0
errOAAI
errJenny Klintman; Niamh Appleby; Basile Stamatopoulos; Katie Ridout; Toby A. Eyre; Pauline Robbe; Laura Lopez Pascua; Samantha J. L. Knight; Helene Dreau; Maite Cabes; Niko Popitsch; Mats Ehinger; Jose I. Martín-Subero; Elías Campo; Robert Månsson; Davide Rossi; Jenny C. Taylor; Dimitrios V. Vavoulis; Anna Schuh
errShare
errSave
researcher View more