arrow
返回

Efficient Stacking and Grasping in Unstructured Environments

delete2024-04-01
delete1
delete
OA
AI
F
Fei Wang *
Y
Yue Liu
M
Manyi Shi
C
Chao Chen
S
Shangdong Liu
J
Jinbiao Zhu
DOI:10.1007/s10846-024-02078-3delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Robotics has been booming in recent years. Especially with the development of artificial intelligence, more and more researchers have devoted themselves to the field of robotics, but there are still many shortcomings in the multi-task operation of robots. Reinforcement learning has achieved good performance in manipulator manipulation, especially in grasping, but grasping is only the first step for the robot to perform actions, and it often ignores the stacking, assembly, placement, and other tasks to be carried out later. Such long-horizon tasks still face the problems of expensive time, dead-end exploration, and process reversal. Hierarchical reinforcement learning has some advantages in solving the above problems, but not all tasks can be learned hierarchically. This paper mainly solves the complex manipulation task of continuous multi-action of the manipulator by improving the method of hierarchical reinforcement learning, aiming to solve the task of long sequences such as stacking and alignment by proposing a framework. Our framework completes simulation experiments on various tasks and improves the success rate from 78.3% to 94.8% when cleaning cluttered toys. In the stacking toy experiment, the training speed is nearly three times faster than the baseline method. And our method can be generalized to other long-horizon tasks. Experiments show that the more complex the task, the greater the advantage of our framework.
Keyword:
Hierarchical reinforcement learning
Long-horizon tasks
Robotic grasp
Stacking

期刊

J
JOURNAL OF INTELLIGENT & ROBOTIC SYSTEMS
IF:
2.8
论文数:
3.9K
被引数:
6.9K

机构

N
northeastern university - china
学者数:
3.1W
论文数: 2.7W
被引数: 37
引用论文

引用论文

Stigma in mental illness: Perspective from eight Asian nations
err2020-01-10
err0
PREAI
errKundadak Ganesh Kudva; Samer El Hayek; Anoop Krishna Gupta; Shunya Kurokawa; Liu Bangshan; Maria Victoria C. Armas‐Villavicencio; Kengo Oishi; Saumya Mishra; Saratcha Tiensuntisook; Norman Sartorius
err分享
err收藏
SWIRL: A sequential windowed inverse reinforcement learning algorithm for robot tasks with delayed rewards
err2018-07-25
err48
PREAI
errKrishnan, Sanjay; Garg, Animesh; Liaw, Richard; Thananjeyan, Brijen; Miller, Lauren; Pokorny, Florian T.; Goldberg, Ken
err分享
err收藏
Poor newborn care practices - a population based survey in eastern Uganda
err2010-02-23
err0
errOAAI
errPeter Waiswa; Stefan Peterson; Goran Tomson; George W Pariyo
err分享
err收藏
Digital Techniques for Wideband Receivers
err
IF0
err2015-12-01
err0
PREAI
errJames Tsui; Chi-Hao Cheng
err分享
err收藏
Uncoupling protein-2 polymorphisms in type 2 diabetes, obesity, and insulin secretion
err2004-01-01
err0
PREAI
errHua Wang; Winston S. Chu; Tong Lu; Sandra J. Hasstedt; Philip A. Kern; Steven C. Elbein
err分享
err收藏
Prediction of the Inhibitory Concentration of Hydroxamic Acids by DFT-QSAR Models on Histone Deacetylase 1
err2018-04-10
err0
PREAI
errDoh Soro; Lynda Ekou; Mamadou Koné; Tchirioua Ekou; Sopi Affi; Lamoussa Ouattara; Nahossé Ziao
err分享
err收藏
Intelligent problem-solving as integrated hierarchical reinforcement learning
err2022-01-25
err40
PREAI
errEppe, Manfred; Gumbsch, Christian; Kerzel, Matthias; Nguyen, Phuong D. H.; Butz, Martin, V; Wermter, Stefan
err分享
err收藏
学者 查看更多内容