arrow
返回

A novel method for learning policies from variable constraint data

delete2009-07-30
delete18
PRE
AI
M
Matthew Howard *
M
Michael Gienger
C
Christian Goerick
S
Sethu Vijayakumar
DOI:10.1007/s10514-009-9129-8delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Many everyday human skills can be framed in terms of performing some task subject to constraints imposed by the environment. Constraints are usually unobservable and frequently change between contexts. In this paper, we present a novel approach for learning (unconstrained) control policies from movement data, where observations come from movements under different constraints. As a key ingredient, we introduce a small but highly effective modification to the standard risk functional, allowing us to make a meaningful comparison between the estimated policy and constrained observations. We demonstrate our approach on systems of varying complexity, including kinematic data from the ASIMO humanoid robot with 27 degrees of freedom, and present results for learning from human demonstration.
Keyword:
Direct policy learning
Constrained motion
Imitation
Nullspace control

期刊

Autonomous Robots 封面图
Autonomous Robots
IF:
4.3
论文数:
1.7K
被引数:
5.0K

机构

H
honda motor company
学者数:
446
论文数: 393
被引数: 0
U
University of Edinburgh
学者数:
5.2W
论文数: 4.6W
被引数: 71
引用论文

引用论文

Natural Actor-Critic
err2008-03-01
err643
PREAI
errPeters, Jan; Schaal, Stefan
err分享
err收藏
Study on the electrical performances of soldered joints between HTS coated-conductors
err2022-03-01
err0
PREAI
errZiyi Huang; Yunfei Tan; Rui He; Yiming Xie; Guangda Wang; Junwen Wei; Yifan Wang; Qiong Wu
err分享
err收藏
Systematic review of efficacy and safety of pemetrexed in non-small-cell-lung cancer
err2014-03-04
err0
PREAI
errMaria Antonia Pérez-Moreno; Mercedes Galván-Banqueri; Sandra Flores-Moreno; Ángela Villalba-Moreno; Jesús Cotrina-Luque; Francisco Javier Bautista-Paloma
err分享
err收藏
The random amplification of polymorphic DNA allows the identification of strains and species of schistosome
err1993-01-01
err0
PREAI
errEmmanuel Dias Neto; Cecilia Pereira de Souza; David Rollinson; Naftale Katz; Sergio D.J. Pena; Andrew J.G. Simpson
err分享
err收藏
Learning to search: Functional gradient techniques for imitation learning
err2009-06-17
err148
PREAI
errRatliff, Nathan D.; Silver, David; Bagnell, J. Andrew
err分享
err收藏
In silico prediction of monovalent and chimeric tetravalent vaccines for prevention and treatment of dengue fever
err2018-01-01
err0
errOAAI
errSubramaniyan Vijayakumar; Venkatachalam Ramesh; Srinivasan Prabhu; Palani Manogar
err分享
err收藏
学者 查看更多内容