arrow
返回

Multisource Transfer Double DQN Based on Actor Learning

delete2018-06-01
delete84
PRE
AI
J
Jie Pan
X
Xuesong Wang *
Y
Yuhu Cheng
Q
Qiang Yu
DOI:10.1109/TNNLS.2018.2806087delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep reinforcement learning (RL) comprehensively uses the psychological mechanisms of trial and error and reward and punishment in RL as well as powerful feature expression and nonlinear mapping in deep learning. Currently, it plays an essential role in the fields of artificial intelligence and machine learning. Since an RL agent needs to constantly interact with its surroundings, the deep Q network (DQN) is inevitably faced with the need to learn numerous network parameters, which results in low learning efficiency. In this paper, a multisource transfer double DQN (MTDDQN) based on actor learning is proposed. The transfer learning technique is integrated with deep RL to make the RL agent collect, summarize, and transfer action knowledge, including policy mimic and feature regression, to the training of related tasks. There exists action overestimation in DQN, i.e., the lower probability limit of action corresponding to the maximum Q value is nonzero. Therefore, the transfer network is trained by using double DQN to eliminate the error accumulation caused by action overestimation. In addition, to avoid negative transfer, i.e., to ensure strong correlations between source and target tasks, a multisource transfer learning mechanism is applied. The Atari2600 game is tested on the arcade learning environment platform to evaluate the feasibility and performance of MTDDQN by comparing it with some mainstream approaches, such as DQN and double DQN. Experiments prove that MTDDQN achieves not only human-like actor learning transfer capability, but also the desired learning efficiency and testing accuracy on target task.
Keyword:
Actor learning
Atari2600 game
double deep Q network (DQN)
multisource transfer
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

暂无机构信息
引用论文

引用论文

Multimorph Eco-Evolutionary Dynamics in Structured Populations
err2022-09-01
err0
errOAAI
errSébastien Lion; Mike Boots; Akira Sasaki
err分享
err收藏
Terrain-Adaptive Locomotion Skills Using Deep Reinforcement Learning
err2016-07-11
err163
PREAI
errBin Peng, Xue; Berseth, Glen; van de Panne, Michiel
err分享
err收藏
err分享
err收藏
学者 查看更多内容