arrow
返回

Large-scale cost function learning for path planning using deep inverse reinforcement learning

delete2017-08-04
delete145
PRE
AI
M
Markus Wulfmeier *
D
Dushyant Rao
I
Ingmar Posner
DOI:10.1177/0278364917722396delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
We present an approach for learning spatial traversability maps for driving in complex, urban environments based on an extensive dataset demonstrating the driving behaviour of human experts. The direct end-to-end mapping from raw input data to cost bypasses the effort of manually designing parts of the pipeline, exploits a large number of data samples, and can be framed additionally to refine handcrafted cost maps produced based on manual hand-engineered features. To achieve this, we introduce a maximum-entropy-based, non-linear inverse reinforcement learning (IRL) framework which exploits the capacity of fully convolutional neural networks (FCNs) to represent the cost model underlying driving behaviours. The application of a high-capacity, deep, parametric approach successfully scales to more complex environments and driving behaviours, while at deployment being run-time independent of training dataset size. After benchmarking against state-of-the-art IRL approaches, we focus on demonstrating scalability and performance on an ambitious dataset collected over the course of 1 year including more than 25,000 demonstration trajectories extracted from over 120 km of urban driving. We evaluate the resulting cost representations by showing the advantages over a carefully, manually designed cost map and furthermore demonstrate its robustness towards systematic errors by learning accurate representations even in the presence of calibration perturbations. Importantly, we demonstrate that a manually designed cost map can be refined to more accurately handle corner cases that are scarcely seen in the environment, such as stairs, slopes and underpasses, by further incorporating human priors into the training framework.
Keyword:
Learning from demonstration
inverse reinforcement learning
neural networks
autonomous driving
cost functions
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

International Journal of Robotics Research 封面图
International Journal of Robotics Research
IF:
5
论文数:
2.4K
被引数:
1.5W

机构

U
university of oxford
学者数:
9.8W
论文数: 8.6W
被引数: 137
引用论文

引用论文

Structural time-varying damage detection using synchrosqueezing wavelet transform
err2015-01-25
err0
PREAI
errJing-Liang Liu; Zuo-Cai Wang; Wei-Xin Ren; Xing-Xin Li
err分享
err收藏
Numerical study of a calamitic liquid-crystal model: Phase behavior and structure
err2005-03-21
err0
PREAI
errGiorgio Cinacchi; Luca De Gaetani; Alessandro Tani
err分享
err收藏
Learning to search: Functional gradient techniques for imitation learning
err2009-06-17
err148
PREAI
errRatliff, Nathan D.; Silver, David; Bagnell, J. Andrew
err分享
err收藏
Co-morbidity in patients with early rheumatoid arthritis - inflammation matters
err2016-01-28
err0
errOAAI
errLena Innala; Clara Sjöberg; Bozena Möller; Lotta Ljung; Torgny Smedby; Anna Södergren; Staffan Magnusson; Solbritt Rantapää-Dahlqvist; Solveig Wållberg-Jonsson
err分享
err收藏
A Multirange Architecture for Collision-Free Off-Road Robot Navigation
err2008-12-17
err29
errOAAI
errSermanet, Pierre; Hadsell, Raia; Scoffier, Marco; Grimes, Matt; Ben, Jan; Erkan, Ayse
err分享
err收藏
Socially compliant mobile robot navigation via inverse reinforcement learning
err2016-07-11
err316
PREAI
errKretzschmar, Henrik; Spies, Markus; Sprunk, Christoph; Burgard, Wolfram
err分享
err收藏
err分享
err收藏
学者 查看更多内容