arrow
返回

Deep Reinforcement Learning: A Survey

delete2024-04-01
delete156
PRE
AI
X
Xu Wang
S
Sen Wang
X
Xingxing Liang
D
Dawei Zhao
J
Jincai Huang
徐鑫 封面图
徐鑫 (Xin Xu)
戴
戴彬 (Bin Dai)
Q
Qiguang Miao *
DOI:10.1109/TNNLS.2022.3207346delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep reinforcement learning (DRL) integrates the feature representation ability of deep learning with the decision-making ability of reinforcement learning so that it can achieve powerful end-to-end learning control capabilities. In the past decade, DRL has made substantial advances in many tasks that require perceiving high-dimensional input and making optimal or near-optimal decisions. However, there are still many challenging problems in the theory and applications of DRL, especially in learning control tasks with limited samples, sparse rewards, and multiple agents. Researchers have proposed various solutions and new theories to solve these problems and promote the development of DRL. In addition, deep learning has stimulated the further development of many subfields of reinforcement learning, such as hierarchical reinforcement learning (HRL), multiagent reinforcement learning, and imitation learning. This article gives a comprehensive overview of the fundamental theories, key algorithms, and primary research domains of DRL. In addition to value-based and policy-based DRL algorithms, the advances in maximum entropy-based DRL are summarized. The future research topics of DRL are also analyzed and discussed.
Keyword:
Task analysis
Mathematical models
Deep learning
Trajectory
Behavioral sciences
Q-learning
Dynamic programming
Deep learning
deep reinforcement learning (DRL)
imitation learning
maximum entropy deep reinforcement learning (RL)
policy gradient
value function

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

X
Xidian University
学者数:
2.4W
论文数: 1.9W
被引数: 9.7K
N
national university of defense technology - china
学者数:
1.8W
论文数: 1.4W
被引数: 9
引用论文

引用论文

Evaluation of iron loading in four types of hepatopancreatic cells of the mangrove crab Ucides cordatus using ferrocene derivatives and iron supplements
err2018-03-27
err0
PREAI
errHector Aguilar Vitorino; Priscila Ortega; Roxana Y. Pastrana Alta; Flavia Pinheiro Zanotto; Breno Pannia Espósito
err分享
err收藏
Anxiety Is Not Associated with the Risk of Dementia or Cognitive Decline: The Rotterdam Study
err2014-12-01
err0
PREAI
errRenée F.A.G. de Bruijn; Nese Direk; Saira Saeed Mirza; Albert Hofman; Peter J. Koudstaal; Henning Tiemeier; M. Arfan Ikram
err分享
err收藏
S1-Leitlinie Post-COVID/Long-COVIDS1-Leitlinie后COVID/Long-COVID
err2021-09-02
err0
errOAAI
errAndreas Rembert Koczulla; Tobias Ankermann; Uta Behrends; Peter Berlit; Sebastian Böing; Folke Brinkmann; Christian Franke; Rainer Glöckl; Christian Gogoll; Thomas Hummel; Juliane Kronsbein; Thomas Maibaum; Eva M. J. Peters; Michael Pfeifer; Thomas Platz; Matthias Pletz; Georg Pongratz; Frank Powitz; Klaus F. Rabe; Carmen Scheibenbogen; Andreas Stallmach; Michael Stegbauer; Hans Otto Wagner; Christiane Waller; Hubert Wirtz; Andreas Zeiher; Ralf Harun Zwick
err分享
err收藏
Learning agile and dynamic motor skills for legged robots
err2019-01-30
err795
errOAAI
errHwangbo, Jemin; Lee, Joonho; Dosovitskiy, Alexey; Bellicoso, Dario; Tsounis, Vassilios; Koltun, Vladlen; Hutter, Marco
err分享
err收藏
err分享
err收藏
学者 查看更多内容