arrow
Return

Model-based inverse reinforcement learning for deterministic systems

delete2022-06-01
delete22
delete
OA
AI
R
Ryan Self *
M
Moad Abudia
S
S M Nahid Mahmud
R
Rushikesh Kamalapurkar
DOI:10.1016/j.automatica.2022.110242delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
This paper focuses on the development of an online data-driven model-based inverse reinforcement learning (MBIRL) technique for linear and nonlinear deterministic systems. Input and output trajectories of an agent under observation, attempting to optimize an unknown reward function, are used to estimate the reward function and the corresponding unknown optimal value function, online and in real-time. To achieve MBIRL using limited data, a novel feedback-driven approach to MBIRL is developed. The feedback policy and the dynamic model of the agent under observation are estimated from the measured data and the estimates are used to generate synthetic data to drive MBIRL. Theoretical guarantees for ultimate boundedness of the estimation errors in general, and convergence of the estimation errors to zero in special cases, are derived using Lyapunov techniques. Proof of concept numerical experiments demonstrates the utility of the developed method to solve linear and nonlinear inverse reinforcement learning problems.(C) 2022 Elsevier Ltd. All rights reserved.
Keywords:
Inverse reinforcement learning
Inverse optimal control
System identification
State estimation
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Automatica cover
Automatica
IF:
5.9
Papers:
1.2W
Citations:
5.2W

Organization

O
oklahoma state university system
Scholars:
8.2K
Papers: 7.3K
Citations: 6