arrow
返回

Optimal path planning approach based on Q-learning algorithm for mobile robots

delete2020-12-01
delete77
PRE
AI
A
Abderraouf Maoudj
A
Abdelfetah Hentout *
DOI:10.1016/j.asoc.2020.106796delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In fact, optimizing path within short computation time still remains a major challenge for mobile robotics applications. In path planning and obstacles avoidance, Q-Learning (QL) algorithm has been widely used as a computational method of learning through environment interaction. However, less emphasis is placed on path optimization using QL because of its slow and weak convergence toward optimal solutions. Therefore, this paper proposes an Efficient Q-Learning (EQL) algorithm to overcome these limitations and ensure an optimal collision-free path in less possible time. In the QL algorithm, successful learning is closely dependent on the design of an effective reward function and an efficient selection strategy for an optimal action that ensures exploration and exploitation. In this regard, a new reward function is proposed to initialize the Q-table and provide the robot with prior knowledge of the environment, followed by a new efficient selection strategy proposal to accelerate the learning process through search space reduction while ensuring a rapid convergence toward an optimized solution. The main idea is to intensify research at each learning stage, around the straight-line segment linking the current position of the robot to Target (optimal path in terms of length). During the learning process, the proposed strategy favors promising actions that not only lead to an optimized path but also accelerate the convergence of the learning process. The proposed EQL algorithm is first validated using benchmarks from the literature, followed by a comparison with other existing QL-based algorithms. The achieved results showed that the proposed EQL gained good learning proficiency; besides, the training performance is significantly improved compared to the state-of-the-art. Concluded, EQL improves the quality of the paths in terms of length, computation time and robot safety, furthermore outperforms other optimization algorithms. (C) 2020 Elsevier B.V. All rights reserved.
Keyword:
Path optimization
Efficient Q-Learning
Efficient selection strategy
Convergence speed
Training performances
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Applied Soft Computing 封面图
Applied Soft Computing
IF:
6.6
论文数:
1.4W
被引数:
4.8W

机构

C
centre for the development of advanced technologies (cdta)
学者数:
132
论文数: 80
被引数: 0
引用论文

引用论文

Reinforcement based mobile robot navigation in dynamic environment
err2011-02-01
err155
PREAI
errJaradat, Mohammad Abdel Kareem; Al-Rousan, Mohammad; Quadan, Lara
err分享
err收藏
Cost-Benefit Analysis
err
IF0
err2007-05-07
err0
PREAI
errEuston Quah; E.J. Mishan; Euston Quah
err分享
err收藏
LEANING OF PROCESSES AND IMPROVING THE WORKING CONDITIONS OF THE NEWLY CREATED WORKING ZONE
err2020-12-31
err0
errOAAI
errVeronika Sabolová; Dagmar Cagáňová; Helena Makyšová
err分享
err收藏
err分享
err收藏
学者 查看更多内容