arrow
Return

Off-Policy Deep Reinforcement Learning Based on Steffensen Value Iteration

delete2021-12-01
delete9
PRE
AI
Y
Yuhu Cheng
L
Lin Chen
陈晨 cover
陈晨 (C. L. Philip Chen)
X
Xuesong Wang *
DOI:10.1109/TCDS.2020.3034452delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
As an important machine learning method, deep reinforcement learning (DRL) has been rapidly developed in recent years and has achieved breakthrough results in many fields, such as video games, natural language processing, and robot control. However, due to the inherit trial-and-error learning mechanism of reinforcement learning and the time-consuming training of deep neural network itself, the convergence speed of DRL is very slow and consequently limits the real applications of DRL. In this article, aiming to improve the convergence speed of DRL, we proposed a novel Steffensen value iteration (SVI) method by applying the Steffensen iteration to the value function iteration of off-policy DRL from the perspective of fixed-point iteration. The proposed SVI is theoretically proved to be convergent and have a faster convergence speed than Bellman value iteration. The proposed SVI has versatility, which can be easily combined with existing off-policy RL algorithms. In this article, we proposed two speedy off-policy DRLs by combining SVI with DDQN and TD3, respectively, namely, SVI-DDQN and SVI-TD3. Experiments on several discrete-action and continuous-action tasks from the Atari 2600 and MuJoCo platforms demonstrated that our proposed SVI-based DRLs can achieve higher average reward in a shorter time than the comparative algorithm.
Keywords:
Reinforcement learning
Acceleration
Convergence
Neural networks
Markov processes
Linear programming
Data mining
Convergence speed
deep reinforcement learning (DRL)
off-policy
Steffensen iteration
value iteration (VI)
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Cognitive and Developmental Systems cover
IEEE Transactions on Cognitive and Developmental Systems
IF:
4.9
Papers:
1.0K
Citations:
3.5K

Organization

U
University of Macau
Scholars:
1.1W
Papers: 1.3W
Citations: 2.0W