arrow
Return

Multi-source transfer ELM-based Q learning

delete2014-08-01
delete31
PRE
AI
J
Jie Pan
X
Xuesong Wang *
Y
Yuhu Cheng
G
G. F. Cao
DOI:10.1016/j.neucom.2013.04.045delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Extreme learning machine (ELM) has advantages of good generalization property, simple structure and convenient calculation. Therefore, an ELM-based Q learning is proposed by using an ELM as a Q-value function approximator, which is suitable for large-scale or continuous space problems. This is the first contribution of this paper. Because the number of ELM hidden layer nodes is equal to that of training samples, large sample size will seriously affect the learning speed. Therefore, a rolling time-window mechanism is introduced into the ELM-based Q learning to reduce the size of training samples of the ELM. In addition, in order to reduce the learning difficulty of new tasks, transfer learning technology is introduced into the ELM-based Q learning. The transfer learning technology can reuse past experience and knowledge to solve current issues. Thus the second contribution is to propose a multi-source transfer ELM-based Q learning (MST-ELMQ), which can take full advantage of valuable information from multiple source tasks and avoid negative transfer resulted from irrelevant information. According to the Bayesian theory, each source task is assigned with a task transfer weight and each source sample is assigned with a sample transfer weight. The task and sample transfer weights determine the number and the manner of transfer samples. Samples with large sample transfer weights are selected from each source task, and assist Q learning agent in quick decision-making for the target task. Simulations results concerning on a boat problem show that MST-ELMQ has better performance than that of Q learning algorithms without or with a single source task, i.e., it can effectively reduce learning difficulty and find an optimal solution with fewer number of training. (C) 2013 Elsevier B.V. All rights reserved.
Keywords:
Q learning
Extreme learning machine
Continuous space
Multi-source transfer
Boat problem

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

No organization information available
Cited Papers

Cited Papers

Animal Models of Myocardial and Vascular Injury
err2010-08-01
err0
PREAI
errAaron M. Abarbanell; Jeremy L. Herrmann; Brent R. Weil; Yue Wang; Jiangning Tan; Steven P. Moberly; Jeremy W. Fiege; Daniel R. Meldrum
errShare
errSave
Robust high performance reinforcement learning through weighted k-nearest neighbors
err2011-03-01
err32
PREAI
errAntonio Martin H, Jose; de Lope, Javier; Maravall, Dario
errShare
errSave
The influence of metallic posts in the detection of vertical root fractures using different imaging examinations
err2014-01-01
err0
errOAAI
errS J M Jakobson; V P D Westphalen; U X Silva Neto; L F Fariniuk; A G D Schroeder; E Carneiro
errShare
errSave
A Survey on Transfer Learning
err2010-10-01
err1.4W
PREAI
errPan, Sinno Jialin; Yang, Qiang
errShare
errSave
Continuous state/action reinforcement learning: A growing self-organizing map approach
err2011-03-01
err18
PREAI
errMontazeri, Hesam; Moradi, Sajjad; Safabakhsh, Reza
errShare
errSave
errShare
errSave
researcher View more