Controlled Markov Chains with AVaR Criteria for Unbounded Costs

Published: 2015/09/27, Updated: 2016/03/29

Applications - OR and Management Sciences average value at risk, markov decision processes, optimal control Short URL: https://optimization-online.org/?p=13641

In this paper, we consider the control problem with the Average-Value-at-Risk (AVaR) criteria of the possibly unbounded $L^{1}$-costs in infinite horizon on a Markov Decision Process (MDP). With a suitable state aggregation and by choosing a priori a global variable $s$ heuristically, we show that there exist optimal policies for the infinite horizon problem. To our knowledge, this is the first work of deriving dynamic programming equations with $L^1$-unbounded costs via AVaR-operator.

Article

Download

View PDF