arrow
Return

Adjustable behavior-guided adaptive dynamic programming for neural learning control

delete2025-07-01
delete0
PRE
AI
G
Guohan Tang
王
王丁 (Ding Wang) *
A
Ao Liu
J
Junfei Qiao
DOI:10.1016/j.neucom.2025.129986delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In this article, an adjustable behavior-guided adaptive dynamic programming (BGADP) algorithm is designed to solve the optimal regulation problem for discrete-time systems. In conventional adaptive dynamic programming methods, gradient information of system dynamics is necessary for conducting policy improvement. However, these methods face challenges when gradient information cannot be computed or when the system dynamics is non-differentiable. To overcome these limitations, a human-behavior-inspired swarm intelligence approach is used to search for superior policies during the iterative process, eliminating the need for gradient information. Additionally, a relaxation factor is introduced into the value function update to accelerate the convergence speed of the algorithm. The monotonicity and convergence properties of the iterative value function are rigorously analyzed. Finally, the effectiveness and practicality of the adjustable BGADP algorithm are validated through two simulation studies, which are implemented using the actor-critic framework with neural networks.
Keywords:
Adaptive dynamic programming
Brain storm optimization
Convergence rate
Neural networks
Optimal control
Swarm intelligence

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

B
Beijing Univ Technol
Scholars:
2.6K
Papers: 1.2K
Citations: 354
Cited Papers

Cited Papers

Advanced value iteration for discrete-time intelligent critic control: A survey
err2023-05-21
err35
PREAI
errZhao, Mingming; Wang, Ding; Qiao, Junfei; Ha, Mingming; Ren, Jin
errShare
errSave
errShare
errSave
Adaptive optimal control for reference tracking independent of exo-system dynamics
err2020-09-01
err7
errOAAI
errKoepf, Florian; Westermann, Johannes; Flad, Michael; Hohmann, Soeren
errShare
errSave
researcher View more