arrow
Return

Scalable Inverse Reinforcement Learning Through Multifidelity Bayesian Optimization

delete2022-08-01
delete51
delete
OA
AI
M
Mahdi Imani *
S
Seyede Fatemeh Ghoreishi
DOI:10.1109/TNNLS.2021.3051012delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Data in many practical problems are acquired according to decisions or actions made by users or experts to achieve specific goals. For instance, policies in the mind of biologists during the intervention process in genomics and metagenomics are often reflected in available data in these domains, or data in cyber-physical systems are often acquired according to actions/decisions made by experts/engineers for purposes, such as control or stabilization. Quantification of experts' policies through available data, which is also known as reward function learning, has been discussed extensively in the literature in the context of inverse reinforcement learning (IRL). However, most of the available techniques come short to deal with practical problems due to the following main reasons: 1) lack of scalability: arising from incapability or poor performance of existing techniques in dealing with large systems and 2) lack of reliability: coming from the incapability of the existing techniques to properly learn the optimal reward function during the learning process. Toward this, in this brief, we propose a multifidelity Bayesian optimization (MFBO) framework that significantly scales the learning process of a wide range of existing IRL techniques. The proposed framework enables the incorporation of multiple approximators and efficiently takes their uncertainty and computational costs into account to balance exploration and exploitation during the learning process. The proposed framework's high performance is demonstrated through genomics, metagenomics, and sets of random simulated problems.
Keywords:
Bayes methods
Computational modeling
Reinforcement learning
Optimization
Linear programming
Uncertainty
Learning systems
Bayesian optimization (BO)
multifidelity models
reinforcement learning (RL)
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Neural Networks and Learning Systems cover
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
Papers:
7.6K
Citations:
7.2W

Organization

G
George Washington University
Scholars:
1.6W
Papers: 1.4W
Citations: 1.7W
University System of Maryland cover
University System of Maryland
Scholars:
6.5W
Papers: 5.6W
Citations: 113
Cited Papers

Cited Papers

Supply Chain Trust
err2019-03-01
err0
PREAI
errNir Kshetri; Jeffrey Voas
errShare
errSave
Towards a wireless and fully-implantable ECoG system
err2013-06-01
err0
PREAI
errE. Tolstosheeva; J. Hoeffmann; J. Pistor; D. Rotermund; T. Schellenberg; D. Boll; T. Hertzberg; V. Gordillo-Gonzalez; S. Mandon; D. Peters-Drolshagen; M. Schneider; K. Pawelzik; A. Kreiter; S. Paul; W. Lang
errShare
errSave
Taking the Human Out of the Loop: A Review of Bayesian Optimization
err2016-01-01
err3.5K
PREAI
errShahriari, Bobak; Swersky, Kevin; Wang, Ziyu; Adams, Ryan P.; de Freitas, Nando
errShare
errSave
Nosema ceranae parasitism impacts olfactory learning and memory and neurochemistry in honey bees (Apis mellifera)
err2017-01-01
err0
errOAAI
errStephanie L. Gage; Catherine Kramer; Samantha Calle; Mark Carroll; Michael Heien; Gloria DeGrandi-Hoffman
errShare
errSave
errShare
errSave
Survival, Fecundity, and Movements of Free‐Roaming Cats
err2010-12-13
err0
PREAI
errPAIGE M. SCHMIDT; ROEL R. LOPEZ; BRET A. COLLIER
errShare
errSave
Hyaluronic Acid Molecular Weight Determines Lung Clearance and Biodistribution after Instillation
err2016-05-24
err0
errOAAI
errChristopher Kuehl; Ti Zhang; Lisa M. Kaminskas; Christopher J. H. Porter; Neal M. Davies; Laird Forrest; Cory Berkland
errShare
errSave
researcher View more