1
Return

SWIRL: A sequential windowed inverse reinforcement learning algorithm for robot tasks with delayed rewards

delete2018-07-25
delete48
PRE
AI
S
Sanjay Krishnan *
A
Animesh Garg
B
Brijen Thananjeyan
F
Florian T. Pokorny
K
Ken Goldberg
DOI:10.1177/0278364918784350delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We present sequential windowed inverse reinforcement learning (SWIRL), a policy search algorithm that is a hybrid of exploration and demonstration paradigms for robot learning. We apply unsupervised learning to a small number of initial expert demonstrations to structure future autonomous exploration. SWIRL approximates a long time horizon task as a sequence of local reward functions and subtask transition conditions. Over this approximation, SWIRL applies Q-learning to compute a policy that maximizes rewards. Experiments suggest that SWIRL requires significantly fewer rollouts than pure reinforcement learning and fewer expert demonstrations than behavioral cloning to learn a policy. We evaluate SWIRL in two simulated control tasks, parallel parking and a two-link pendulum. On the parallel parking task, SWIRL achieves the maximum reward on the task with 85% fewer rollouts than Q-learning, and one-eight of demonstrations needed by behavioral cloning. We also consider physical experiments on surgical tensioning and cutting deformable sheets using a da Vinci surgical robot. On the deformable tensioning task, SWIRL achieves a 36% relative improvement in reward compared with a baseline of behavioral cloning with segmentation.
Keywords:
Reinforcement learning
inverse reinforcement learning
learning from demonstrations
medical robots and systems
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

International Journal of Robotics Research cover
International Journal of Robotics Research
IF:
5
Papers:
2.4K
Citations:
1.5W

Organization

S
Stanford University
Scholars:
9.6W
Papers: 8.2W
Citations: 17.0W
University of California System cover
University of California System
Scholars:
37.2W
Papers: 33.6W
Citations: 6.6K
Cited Papers

Cited Papers

Citing Papers

Citing Papers