arrow
Return

(Approximate) iterated successive approximations algorithm for sequential decision processes

delete2012-02-08
delete3
PRE
AI
P
Pelin G. Canbolat
U
Uriel G. Rothblum *
DOI:10.1007/s10479-012-1073-xdelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The paper proves the convergence of (Approximate) Iterated Successive Approximations Algorithm for solving infinite-horizon sequential decision processes satisfying the monotone contraction assumption. At every stage of this algorithm, the value function at hand is used as a terminal reward to determine an (approximately) optimal policy for the one-period problem. This policy is then iterated for a (finite or infinite) number of times and the resulting return function is used as the starting value function for the next stage of the scheme. This method generalizes the standard successive approximations, policy iteration and Denardo's generalization of the latter.
Keywords:
Sequential decision processes
Markov decision chains
Successive approximations
Modified policy iteration

Journal

Annals of Operations Research cover
Annals of Operations Research
IF:
4.5
Papers:
8.0K
Citations:
2.1W

Organization

T
Technion Israel Institute of Technology
Scholars:
1.6W
Papers: 1.5W
Citations: 2.0W
Cited Papers

Cited Papers

Coconut allergy
err2021-05-01
err0
errOAAI
errLacey Kruse; Jennifer Lor; Rame Yousif; Jacqueline A. Pongracic; Anna B. Fishbein
errShare
errSave
errShare
errSave
researcher View more