arrow
Return

Active reward learning with a novel acquisition function

delete2015-07-16
delete39
PRE
AI
C
Christian Daniel *
O
Oliver Kroemer
M
Malte Viering
J
Jan Metz
J
Jan Peters
DOI:10.1007/s10514-015-9454-zdelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Reward functions are an essential component of many robot learning methods. Defining such functions, however, remains hard in many practical applications. For tasks such as grasping, there are no reliable success measures available. Defining reward functions by hand requires extensive task knowledge and often leads to undesired emergent behavior. We introduce a framework, wherein the robot simultaneously learns an action policy and a model of the reward function by actively querying a human expert for ratings. We represent the reward model using a Gaussian process and evaluate several classical acquisition functions (AFs) from the Bayesian optimization literature in this context. Furthermore, we present a novel AF, expected policy divergence. We demonstrate results of our method for a robot grasping task and show that the learned reward function generalizes to a similar task. Additionally, we evaluate the proposed novel AF on a real robot pendulum swing-up task.
Keywords:
Reinforcement learning
Active learning
Bayesian optimization
Preference learning
Inverse reinforcement learning
Reward functions
Acquisition functions
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Autonomous Robots cover
Autonomous Robots
IF:
4.3
Papers:
1.7K
Citations:
5.0K

Organization

T
Technical University of Darmstadt
Scholars:
1.3W
Papers: 10.0K
Citations: 1.2W