arrow
返回

Shrinking-horizon dynamic programming

delete2010-02-01
delete31
delete
OA
AI
J
Joëlle Skaf *
S
Stephen Boyd
A
Assaf Zeevi
DOI:10.1002/rnc.1566delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
We describe a heuristic control policy for a general finite-horizon stochastic control problem, which can be used when the current process disturbance is not conditionally independent of the previous disturbances, given the current state. At each time step, we approximate the distribution of future disturbances (conditioned on what has been observed) by a product distribution with the same marginals. We then carry out dynamic programming (DP), using this modified future disturbance distribution, to find an optimal policy, and in particular, the optimal current action. We then execute only the optimal current action. At the next step, we update the conditional distribution, and repeat the process, this time with a horizon reduced by one step. (This explains the name 'shrinking-horizon dynamic programming'). We explain how the method can be thought of as an extension of model predictive control, and illustrate our method on two variations on a revenue management problem. Copyright (C) 2010 John Wiley & Sons, Ltd.
Keyword:
dynamic programming
model predictive control
revenue management

期刊

International Journal of Robust and Nonlinear Control 封面图
International Journal of Robust and Nonlinear Control
IF:
3.2
论文数:
7.0K
被引数:
1.4W

机构

A
alphabet inc.
学者数:
1.1K
论文数: 663
被引数: 0
S
Stanford University
学者数:
9.6W
论文数: 8.2W
被引数: 17.0W
G
Google Incorporated
学者数:
3.5K
论文数: 1.8K
被引数: 8
学者 查看更多机构