arrow
返回

Dynamic sparse coding-based value estimation network for deep reinforcement learning

delete2023-11-01
delete4
PRE
AI
赵昊立 封面图
赵昊立 (Zhao, Haoli)
Z
Zhenni Li *
W
Wensheng Su
Xie Shengli 封面图
Xie Shengli (Shengli Xie)
DOI:10.1016/j.neunet.2023.09.013delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep Reinforcement Learning (DRL) is one powerful tool for varied control automation problems. Performances of DRL highly depend on the accuracy of value estimation for states from environments. However, the Value Estimation Network (VEN) in DRL can be easily influenced by the phenomenon of catastrophic interference from environments and training. In this paper, we propose a Dynamic Sparse Coding-based (DSC) VEN model to obtain precise sparse representations for accurate value prediction and sparse parameters for efficient training, which is not only applicable in Q-learning structured discrete-action DRL but also in actor-critic structured continuous-action DRL. In detail, to alleviate interference in VEN, we propose to employ DSC to learn sparse representations for accurate value estimation with dynamic gradients beyond the conventional l1 norm that provides same-value gradients. To avoid influences from redundant parameters, we employ DSC to prune weights with dynamic thresholds more efficiently than static thresholds like l1 norm. Experiments demonstrate that the proposed algorithms with dynamic sparse coding can obtain higher control performances than existing benchmark DRL algorithms in both discrete-action and continuous-action environments, e.g., over 25% increase in Puddle World and about 10% increase in Hopper. Moreover, the proposed algorithm can reach convergence efficiently with fewer episodes in different environments.(c) 2023 Elsevier Ltd. All rights reserved.
Keyword:
Deep reinforcement learning
Value estimation network
Dynamic sparse coding

期刊

Neural Networks 封面图
Neural Networks
IF:
6.3
论文数:
8.2K
被引数:
3.0W

机构

G
guangdong university of technology
学者数:
3.0W
论文数: 2.0W
被引数: 36
引用论文

引用论文

Deep reinforcement learning for wireless sensor scheduling in cyber-physical systems
err2020-03-01
err92
errOAAI
errLeong, Alex S.; Ramaswamy, Arunselvan; Quevedo, Daniel E.; Karl, Holger; Shi, Ling
err分享
err收藏
Deep reinforcement learning guided graph neural networks for brain network analysis用于脑网络分析的深度强化学习引导图神经网络
err2022-10-01
err33
errOAAI
errZhao, Xusheng; Wu, Jia; Peng, Hao; Beheshti, Amin; Monaghan, Jessica J. M.; McAlpine, David; Hernandez-Perez, Heivet; Dras, Mark; Dai, Qiong; Li, Yangyang; Yu, Philip S.; He, Lifang
err分享
err收藏
A novel deep policy gradient action quantization for trusted collaborative computation in intelligent vehicle networks
err2023-07-01
err16
PREAI
errChen, Miaojiang; Yi, Meng; Huang, Mingfeng; Huang, Guosheng; Ren, Yingying; Liu, Anfeng
err分享
err收藏
The intelligent critic framework for advanced optimal control
err2022-01-16
err129
PREAI
errWang, Ding; Ha, Mingming; Zhao, Mingming
err分享
err收藏
学者 查看更多内容