An Approximate Stochastic Annealing algorithm for finite horizon Markov decision processes

Jiaqiao Hu,Hyeong Soo Chang

doi:10.1109/cdc.2010.5717689

An Approximate Stochastic Annealing algorithm for finite horizon Markov decision processes

Jiaqiao Hu, Hyeong Soo Chang

https://doi.org/10.1109/cdc.2010.5717689

Copy DOI

Publication Date: Dec 1, 2010

Citations: 16

Affiliation: State University of New York, Stony Brook University, Sogang University

#Finite Horizon Markov Decision Processes #Markov Decision Processes + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

We present a simulation-based algorithm called Approximate Stochastic Annealing (ASA) for solving finite-horizon Markov decision processes (MDPs). The algorithm iteratively estimates the optimal policy by sampling from a sequence of probability distribution functions over the policy space. By exploiting a novel connection of ASA to the stochastic approximation method, we show that the sequence of distribution functions generated by the algorithm converges to a degenerated distribution that concentrates only on the optimal policy. Numerical examples are also provided to illustrate the algorithm.

Full Text