arrow
返回

Markov decision processes with recursive risk measures

delete2022-02-01
delete14
delete
OA
AI
N
Nicole Bäuerle *
A
Alexander Glauner
DOI:10.1016/j.ejor.2021.04.030delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In this paper, we consider risk-sensitive Markov Decision Processes (MDPs) with Borel state and action spaces and unbounded cost. We treat both finite and infinite planning horizons. Our optimality criterion is based on the recursive application of static risk measures. This is motivated by recursive utilities in the economic literature. It has been studied before for the entropic risk measure and is extended here to general static risk measures. Under direct assumptions on the model data we derive a Bellman equa-tion and prove the existence of optimal Markov policies. For an infinite planning horizon, the model is shown to be contractive and the optimal policy to be stationary. Our approach unifies results for a num-ber of well-known risk measures. Moreover, we establish a connection to distributionally robust MDPs, which provides a global interpretation of the recursively defined objective function. Monotone models are studied in particular. (c) 2021 Elsevier B.V. All rights reserved.
Keyword:
Dynamic programming
Risk-sensitive Markov decision process
Risk measure
Robustness
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

European Journal of Operational Research 封面图
European Journal of Operational Research
IF:
6
论文数:
2.2W
被引数:
6.4W

机构

H
Helmholtz Association
学者数:
13.2W
论文数: 10.7W
被引数: 145
引用论文

引用论文

Shotgun DNA microarrays and stage‐specific gene expression in Plasmodium falciparum malaria
err2002-03-01
err0
errOAAI
errRhian E. Hayward; Joseph L. DeRisi; Suad Alfadhli; David C. Kaslow; Patrick O. Brown; Pradipsinh K. Rathod
err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容