Return
PDP: Parallel Dynamic Programming
DOI:10.1109/JAS.2017.7510310.png)
Abstract
En 中文
Deep reinforcement learning is a focus research area in artificial intelligence. The principle of optimality in dynamic programming is a key to the success of reinforcement learning methods. The principle of adaptive dynamic programming (ADP) is first presented instead of direct dynamic programming (DP), and the inherent relationship between ADP and deep reinforcement learning is developed. Next, analytics intelligence, as the necessary requirement, for the real reinforcement learning, is discussed. Finally, the principle of the parallel dynamic programming, which integrates dynamic programming and analytics intelligence, is presented as the future computational intelligence.
Keywords:
Parallel dynamic programming
Dynamic programming
Adaptive dynamic programming
Reinforcement learning
Deep learning
Neural networks
Artificial intelligence
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
I
IF:
19.2
Papers:
1.4K
Citations:
1.1W

