arrow
Return

Continuous probabilistic model building genetic network programming using reinforcement learning

delete2015-02-01
delete5
PRE
AI
X
Xianneng Li
K
Kotaro Hirasawa *
DOI:10.1016/j.asoc.2014.10.023delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Recently, a novel probabilistic model-building evolutionary algorithm (so called estimation of distribution algorithm, or EDA), named probabilistic model building genetic network programming (PMBGNP), has been proposed. PMBGNP uses graph structures for its individual representation, which shows higher expression ability than the classical EDAs. Hence, it extends EDAs to solve a range of problems, suchas data mining and agent control. This paper is dedicated to propose a continuous version of PMBGNP for continuous optimization in agent control problems. Different from the other continuous EDAs, the proposed algorithm evolves the continuous variables by reinforcement learning (RL). We compare the performance with several state-of-the-art algorithms on a real mobile robot control problem. The results show that the proposed algorithm outperforms the others with statistically significant differences. (C) 2014 Elsevier B.V. All rights reserved.
Keywords:
Estimation of distribution algorithm
Probabilistic model building genetic network programming
Continuous optimization
Reinforcement learning
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Applied Soft Computing cover
Applied Soft Computing
IF:
6.6
Papers:
1.4W
Citations:
4.8W

Organization

W
Waseda University
Scholars:
1.0W
Papers: 8.7K
Citations: 8.3K
Cited Papers

Cited Papers

errShare
errSave
Compact Differential Evolution
err2011-02-01
err200
PREAI
errMininno, Ernesto; Neri, Ferrante; Cupertino, Francesco; Naso, David
errShare
errSave
Capillary blood flow in patients with dysmenorrhea treated with acupuncture
err2013-12-01
err0
errOAAI
errTao Huang; Lijian Yang; Shuyong Jia; Xiang Mu; Mozheng Wu; Hang Ye; Weizhe Liu; Xinnong Cheng
errShare
errSave
err
IF0
err
err0
PREAI
err
errShare
errSave
researcher View more