arrow
返回

Visual Navigation With Multiple Goals Based on Deep Reinforcement Learning

delete2021-12-01
delete31
PRE
AI
Z
Zhenhuan Rao
Y
Yuechen Wu
Z
Zifei Yang
W
Wei Zhang *
LU Shijian 封面图
LU Shijian (Shijian Lu)
W
Weizhi Lu
Z
Zheng-Jun Zha
DOI:10.1109/TNNLS.2021.3057424delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Learning to adapt to a series of different goals in visual navigation is challenging. In this work, we present a model-embedded actor-critic architecture for the multigoal visual navigation task. To enhance the task cooperation in multigoal learning, we introduce two new designs to the reinforcement learning scheme: inverse dynamics model (InvDM) and multigoal colearning (MgCl). Specifically, InvDM is proposed to capture the navigation-relevant association between state and goal and provide additional training signals to relieve the sparse reward issue. MgCl aims at improving the sample efficiency and supports the agent to learn from unintentional positive experiences. Besides, to further improve the scene generalization capability of the agent, we present an enhanced navigation model that consists of two self-supervised auxiliary task modules. The first module, which is named path closed-loop detection, helps to understand whether the state has been experienced. The second one, namely the state-target matching module, tries to figure out the difference between state and goal. Extensive results on the interactive platform AI2-THOR demonstrate that the agent trained with the proposed method converges faster than state-of-the-art methods while owning good generalization capability. The video demonstration is available at https://vsislab.github.io/mgvn.
Keyword:
Navigation
Task analysis
Visualization
Training
Reinforcement learning
Computer architecture
Adaptation models
Deep reinforcement learning
scene generalization
visual navigation
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

N
Nanyang Technological University
学者数:
4.9W
论文数: 4.8W
被引数: 8.1W
S
shandong university
学者数:
9.5W
论文数: 6.4W
被引数: 94
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
学者 查看更多机构
引用论文

引用论文

Machine grading of lumber : practical concerns for lumber producers
err
IF0
err2000-01-01
err0
PREAI
errWilliam L. Galligan; Kent A. McDonald
err分享
err收藏
White matter integrity in major depressive disorder: Implications of childhood trauma, 5-HTTLPR and BDNF polymorphisms
err2016-07-01
err0
PREAI
errErica L. Tatham; Rajamannar Ramasubbu; Ismael Gaxiola-Valdez; Filomeno Cortese; Darren Clark; Bradley Goodyear; Jane Foster; Geoffrey B. Hall
err分享
err收藏
err分享
err收藏
Why ResNet Works? Residuals Generalize
err2020-12-01
err173
errOAAI
errHe, Fengxiang; Liu, Tongliang; Tao, Dacheng
err分享
err收藏
err分享
err收藏
学者 查看更多内容