arrow
Return

Feature-based methods for large scale dynamic programming

delete1996-03-01
delete334
delete
OA
AI
J
John N. Tsitsiklis *
V
VanRoy, B
DOI:10.1007/BF00114724delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We develop a methodological framework and present a few different ways in which dynamic programming and compact representations can be combined to solve large scale stochastic control problems. In particular, we develop algorithms that employ two types of feature-based compact representations; that is, representations that involve feature extraction and a relatively simple approximation architecture. We prove the convergence of these algorithms and provide bounds on the approximation error. As an example, one of these algorithms is used to generate a strategy for the game of Tetris. Furthermore, we provide a counterexample illustrating the difficulties of integrating compact representations with dynamic programming, which exemplifies the shortcomings of certain simple approaches.
Keywords:
compact representation
curse of dimensionality
dynamic programming
features
function approximation
neuro-dynamic programming
reinforcement learning

Journal

Machine Learning cover
Machine Learning
IF:
2.9
Papers:
2.7K
Citations:
3.4W

Organization

No organization information available
Cited Papers

Cited Papers

TRIPPD: A Practice-Based Network Effectiveness Study of Postpartum Depression Screening and Management
err2012-07-09
err0
errOAAI
errB. P. Yawn; A. J. Dietrich; P. Wollan; S. Bertram; D. Graham; J. Huff; M. Kurland; S. Madison; W. D. Pace
errShare
errSave
err
IF0
err
err0
errOAAI
err
errShare
errSave
NETWORKS FOR APPROXIMATION AND LEARNING
err1990-01-01
err2.1K
PREAI
errPOGGIO, T; GIROSI, F
errShare
errSave
no more