arrow
返回

Approximate dynamic programming via iterated Bellman inequalities

delete2014-02-19
delete42
delete
OA
AI
王
王阳 (Yan Wang) *
B
Brendan O’Donoghue
S
Stephen Boyd
DOI:10.1002/rnc.3152delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this paper, we introduce new methods for finding functions that lower bound the value function of a stochastic control problem, using an iterated form of the Bellman inequality. Our method is based on solving linear or semidefinite programs, and produces both a bound on the optimal objective, as well as a suboptimal policy that appears to works very well. These results extend and improve bounds obtained in a previous paper using a single Bellman inequality condition. We describe the methods in a general setting and show how they can be applied in specific cases including the finite state case, constrained linear quadratic control, switched affine control, and multi-period portfolio investment. Copyright (c) 2014 John Wiley & Sons, Ltd.
Keyword:
convex optimization
dynamic programming
stochastic control

期刊

International Journal of Robust and Nonlinear Control 封面图
International Journal of Robust and Nonlinear Control
IF:
3.2
论文数:
7.0K
被引数:
1.4W

机构

S
Stanford University
学者数:
9.6W
论文数: 8.2W
被引数: 17.0W
引用论文

引用论文

Development and evaluation of a field-based high-throughput phenotyping platform
err2014-01-01
err0
errOAAI
errPedro Andrade-Sanchez; Michael A. Gore; John T. Heun; Kelly R. Thorp; A. Elizabete Carmo-Silva; Andrew N. French; Michael E. Salvucci; Jeffrey W. White
err分享
err收藏
TRIPPD: A Practice-Based Network Effectiveness Study of Postpartum Depression Screening and Management
err2012-07-09
err0
errOAAI
errB. P. Yawn; A. J. Dietrich; P. Wollan; S. Bertram; D. Graham; J. Huff; M. Kurland; S. Madison; W. D. Pace
err分享
err收藏
Fast computation of optimal contact forces
err2007-12-01
err98
errOAAI
errBoyd, Stephen P.; Wegbreit, Ben
err分享
err收藏
Relaxing dynamic programming
err2006-08-01
err260
PREAI
errLincoln, Bo; Rantzer, Anders
err分享
err收藏
err分享
err收藏
学者 查看更多内容