arrow
返回

Iterative adaptive dynamic programming methods with neural network implementation for multi-player zero-sum games

delete2018-09-01
delete33
PRE
AI
姜贺 封面图
姜贺 (He Jiang)
张化光 封面图
张化光 (Huaguang Zhang) *
J
Ji Han
K
Kun Zhang
DOI:10.1016/j.neucom.2018.04.005delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This paper presents novel iterative learning methods along with the neural network implementation for multi-player zero-sum games. Solving zero-sum games depends on the solutions of Hamilton-Jacobi-Isaacs equations, which are nonlinear partial differential equations. These solutions are generally difficult or even impossible to be obtained analytically. To overcome this difficulty, iterative adaptive dynamic programming algorithms are utilized. In the related research works, three-network architecture, i.e., critic-actor-disturbance structure, is used to approximate the value function, control policies and disturbance policies. Different from the previous works, this paper employs single-network architecture, i.e., critic-only structure, to implement the proposed algorithms, which reduces the computation burden and the complexity of design procedure. Finally, two simulation examples are provided to illustrate the effectiveness of our proposed methods. (C) 2018 Published by Elsevier B.V.
Keyword:
Adaptive dynamic programming
Approximate dynamic programming
Zero-sum games
Neural networks
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

N
northeastern university - china
学者数:
3.1W
论文数: 2.7W
被引数: 37
引用论文

引用论文

Spectral Analysis of Epidemic Thresholds of Temporal Networks
err2020-05-01
err36
PREAI
errZhang, Yi-Qing; Li, Xiang; Vasilakos, Athanasios V.
err分享
err收藏
学者 查看更多内容