arrow
返回

Scalable Inverse Reinforcement Learning Through Multifidelity Bayesian Optimization

delete2022-08-01
delete51
delete
OA
AI
M
Mahdi Imani *
S
Seyede Fatemeh Ghoreishi
DOI:10.1109/TNNLS.2021.3051012delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Data in many practical problems are acquired according to decisions or actions made by users or experts to achieve specific goals. For instance, policies in the mind of biologists during the intervention process in genomics and metagenomics are often reflected in available data in these domains, or data in cyber-physical systems are often acquired according to actions/decisions made by experts/engineers for purposes, such as control or stabilization. Quantification of experts' policies through available data, which is also known as reward function learning, has been discussed extensively in the literature in the context of inverse reinforcement learning (IRL). However, most of the available techniques come short to deal with practical problems due to the following main reasons: 1) lack of scalability: arising from incapability or poor performance of existing techniques in dealing with large systems and 2) lack of reliability: coming from the incapability of the existing techniques to properly learn the optimal reward function during the learning process. Toward this, in this brief, we propose a multifidelity Bayesian optimization (MFBO) framework that significantly scales the learning process of a wide range of existing IRL techniques. The proposed framework enables the incorporation of multiple approximators and efficiently takes their uncertainty and computational costs into account to balance exploration and exploitation during the learning process. The proposed framework's high performance is demonstrated through genomics, metagenomics, and sets of random simulated problems.
Keyword:
Bayes methods
Computational modeling
Reinforcement learning
Optimization
Linear programming
Uncertainty
Learning systems
Bayesian optimization (BO)
multifidelity models
reinforcement learning (RL)
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

G
George Washington University
学者数:
1.6W
论文数: 1.4W
被引数: 1.7W
University System of Maryland 封面图
University System of Maryland
学者数:
6.5W
论文数: 5.6W
被引数: 113
引用论文

引用论文

Supply Chain Trust
err2019-03-01
err0
PREAI
errNir Kshetri; Jeffrey Voas
err分享
err收藏
Towards a wireless and fully-implantable ECoG system
err2013-06-01
err0
PREAI
errE. Tolstosheeva; J. Hoeffmann; J. Pistor; D. Rotermund; T. Schellenberg; D. Boll; T. Hertzberg; V. Gordillo-Gonzalez; S. Mandon; D. Peters-Drolshagen; M. Schneider; K. Pawelzik; A. Kreiter; S. Paul; W. Lang
err分享
err收藏
Taking the Human Out of the Loop: A Review of Bayesian Optimization将人类带出循环: 贝叶斯优化的回顾
err2016-01-01
err3.5K
PREAI
errShahriari, Bobak; Swersky, Kevin; Wang, Ziyu; Adams, Ryan P.; de Freitas, Nando
err分享
err收藏
Nosema ceranae parasitism impacts olfactory learning and memory and neurochemistry in honey bees (Apis mellifera)
err2017-01-01
err0
errOAAI
errStephanie L. Gage; Catherine Kramer; Samantha Calle; Mark Carroll; Michael Heien; Gloria DeGrandi-Hoffman
err分享
err收藏
err分享
err收藏
Survival, Fecundity, and Movements of Free‐Roaming Cats
err2010-12-13
err0
PREAI
errPAIGE M. SCHMIDT; ROEL R. LOPEZ; BRET A. COLLIER
err分享
err收藏
Hyaluronic Acid Molecular Weight Determines Lung Clearance and Biodistribution after Instillation
err2016-05-24
err0
errOAAI
errChristopher Kuehl; Ti Zhang; Lisa M. Kaminskas; Christopher J. H. Porter; Neal M. Davies; Laird Forrest; Cory Berkland
err分享
err收藏
学者 查看更多内容