A variance minimization problem for a Markov decision process

Hajime Kawai

doi:10.1016/0377-2217(87)90148-2

A variance minimization problem for a Markov decision process

Hajime Kawai

https://doi.org/10.1016/0377-2217(87)90148-2

Copy DOI

Journal: European Journal of Operational Research	Publication Date: Jul 1, 1987
Citations: 34

Affiliation: Osaka University

#Discrete Time Markov Decision Process #Markov Decision Process + Show 8 more

Abstract
Full-Text
Similar Papers

Abstract

In the steady state of a discrete time Markov decision process, we consider the problem to find an optimal randomized policy that minimizes the variance of the reward in a transition among the policies which give the mean not less than a specified value. The problem is solved by introducing a parametric Markov decision process with average cost criterion. It is shown that there exists an optimal policy which is a mixture of at most two pure policies. As an application, the toymaker's problem is discussed.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: European Journal of Operational Research

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.