arrow
返回

Reinforcement learning method based on sample regularization and adaptive learning rate for AGV path planning

delete2025-01-01
delete0
PRE
AI
J
Jun Nie
G
Guihua Zhang
L
Lu Xiao
王海霞 封面图
王海霞 (Haixia Wang) *
C
Chunyang Sheng
L
L. J. Sun
DOI:10.1016/j.neucom.2024.128820delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This paper proposes the proximal policy optimization (PPO) method based on sample regularization (SR) and adaptive learning rate (ALR) to address the issues of limited exploration ability and slow convergence speed in Autonomous Guided Vehicle (AGV) path planning using reinforcement learning algorithms in dynamic environments. Firstly, the regularization term based on empirical samples is designed to solve the bias and imbalance issues of training samples, and the sample regularization is added to the objective function to improve the policy selectivity of the PPO algorithm, thereby increasing the AGV's exploration ability during the training process in the working environment. Secondly, the Fisher information matrix of the Kullback-Leibler (KL) divergence approximation and the KL divergence constraint term are exploited to design the policy update mechanism based on the dynamically adjustable adaptive learning rate throughout training. The method considers the geometric structure of the parameter space and the change of the policy gradient, aiming to optimize parameter update direction and enhance convergence speed and stability of the algorithm. Finally, the AGV path planning scheme based on reinforcement learning is established for simulation verification and comparations in two-dimensional raster map and Gazebo 3D simulation environment. Simulation results verify the feasibility and superiority of the proposed method applied to the AGV path planning problem.
Keyword:
Reinforcement learning
AGV
Path planning
Sample regularization
Adaptive learning rate

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

暂无机构信息
引用论文

引用论文

err分享
err收藏
err分享
err收藏
err分享
err收藏
Binary nanosheet frameworks of graphene/polyaniline composite for high-areal flexible supercapacitors
err2021-11-01
err0
PREAI
errFeng Shao; Yaqiong Niu; Bin Li; Gang Li; Zhi Yang; Yanjie Su; Yafei Zhang; Nantao Hu
err分享
err收藏
学者 查看更多内容