arrow
Return

A novel neural reinforcement learning algorithm based on critic-only architecture for continuous state control problems

delete2025-10-31
delete0
PRE
AI
O
Omid Mehrabi
A
Ahmad Fakharian *
M
Mehdi Siahi
A
Amin Ramezani
DOI:10.1007/s00500-025-10875-7delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Many real-world control problems inherently have large and continuous domains, leading to the curse of dimensionality in the learning process. This paper proposes a novel critic-only based Neural Reinforcement Learning (NRL) algorithm designed for continuous state spaces. Our approach, called Neural Least Square Policy Iteration (NLSPI), utilizes LSPI method with Radial Basis Function (RBF) network as a function approximator. RBF network offers a continuous and compact representation for continues sensory spaces, enabling the learning system to generalize the learned policy to unseen states. LSPI is employed to adjust the weight parameters of RBF network. Additionally, we present positive theoretical results regarding an error bound between the optimal and the approximated Action Value Function (AVF) for NLSPI. Our proposed method boasts favorable features such as positive mathematical analysis, independence from learning rate, and comparatively good convergence properties. Simulation studies demonstrate the applicability and performance of our learning framework. The overall results indicate that the proposed idea can outperform previously known reinforcement learning algorithms.
Keywords:
Neural reinforcement learning
RBF network
Generalization
Least square policy iteration

Journal

Soft Computing cover
Soft Computing
IF:
2.5
Papers:
1.0W
Citations:
2.1W

Organization

D
department of electrical engineering
Scholars:
1.3K
Papers: 671
Citations: 0