Computational Methods for Risk-Averse Undiscounted Transient Markov Models

Özlem Çavuş,Andrzej Ruszczyński

doi:10.1287/opre.2013.1251

Computational Methods for Risk-Averse Undiscounted Transient Markov Models

Özlem Çavuş, Andrzej Ruszczyński

Open Access

https://doi.org/10.1287/opre.2013.1251

Copy DOI

Journal: Operations Research	Publication Date: Apr 1, 2014
Citations: 33

Affiliation: Bilkent University, Rutgers, The State University of New Jersey

#Nonsmooth Newton Method #Policy Iteration Method + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

The total cost problem for discrete-time controlled transient Markov models is considered. The objective functional is a Markov dynamic risk measure of the total cost. Two solution methods, value and policy iteration, are proposed, and their convergence is analyzed. In the policy iteration method, we propose two algorithms for policy evaluation: the nonsmooth Newton method and convex programming, and we prove their convergence. The results are illustrated on a credit limit control problem.

Full Text