arrow
返回

Feature-based methods for large scale dynamic programming

delete1996-03-01
delete334
delete
OA
AI
J
John N. Tsitsiklis *
V
VanRoy, B
DOI:10.1007/BF00114724delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
We develop a methodological framework and present a few different ways in which dynamic programming and compact representations can be combined to solve large scale stochastic control problems. In particular, we develop algorithms that employ two types of feature-based compact representations; that is, representations that involve feature extraction and a relatively simple approximation architecture. We prove the convergence of these algorithms and provide bounds on the approximation error. As an example, one of these algorithms is used to generate a strategy for the game of Tetris. Furthermore, we provide a counterexample illustrating the difficulties of integrating compact representations with dynamic programming, which exemplifies the shortcomings of certain simple approaches.
Keyword:
compact representation
curse of dimensionality
dynamic programming
features
function approximation
neuro-dynamic programming
reinforcement learning

期刊

Machine Learning 封面图
Machine Learning
IF:
2.9
论文数:
2.7K
被引数:
3.4W

机构

暂无机构信息
引用论文

引用论文

TRIPPD: A Practice-Based Network Effectiveness Study of Postpartum Depression Screening and Management
err2012-07-09
err0
errOAAI
errB. P. Yawn; A. J. Dietrich; P. Wollan; S. Bertram; D. Graham; J. Huff; M. Kurland; S. Madison; W. D. Pace
err分享
err收藏
err
IF0
err
err0
errOAAI
err
err分享
err收藏
err分享
err收藏
NETWORKS FOR APPROXIMATION AND LEARNING
err1990-01-01
err2.1K
PREAI
errPOGGIO, T; GIROSI, F
err分享
err收藏
没有更多内容