arrow
Return

An Approximate Quadratic Programming for Efficient Bellman Equation Solution

delete2019-01-01
delete1
delete
OA
AI
J
Jianmei Su
H
Hong Cheng *
H
Hongliang Guo
R
Rui Huang
Z
Zhinan Peng
DOI:10.1109/ACCESS.2019.2939161delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
This paper proposes an efficient algorithm which relies on quadratic programming for approximately solving the Bellman equation in reinforcement learning problem and guarantees to return optimal decision parameters. Through further applying universal approximation and fixed cardinality minimization techniques, the proposed algorithm in one hand expands the representation ability of basic linear value functions, on the other hand, it guarantees the convergence of the Bellman error. Experimental results on two canonical reinforcement learning scenarios demonstrate that the proposed algorithm achieves similar or better performance than the state-of-the-art algorithms, while reduces the computation time significantly and improves the robustness of the algorithm against state uncertainty.
Keywords:
Markov decision processes
approximate quadratic programming
Bellman equation solutions
universal approximation
fixed cardinality
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Access cover
IEEE Access
IF:
3.6
Papers:
9.8W
Citations:
29.4W

Organization

No organization information available
Cited Papers

Cited Papers

Qipengyuania soli sp. nov., Isolated from Mangrove Soil
err2021-05-28
err0
PREAI
errYang Liu; Tao Pei; Ming-Rong Deng; Honghui Zhu
errShare
errSave
Introduction
err2005-06-01
err0
PREAI
errD. F. Barbe
errShare
errSave
Efficient approximate linear programming for factored MDPs
err2015-08-01
err6
errOAAI
errChen, Feng; Cheng, Qiang; Dong, Jianwu; Yu, Zhaofei; Wang, Guojun; Xu, Wenli
errShare
errSave
researcher View more