arXiv · 1512.00583
Central-limit approach to risk-aware Markov decision processes
Abstract
Whereas classical Markov decision processes maximize the expected reward, we consider minimizing the risk. We propose to evaluate the risk associated to a given policy over a long-enough time horizon with the help of a central limit theorem. The proposed approach works whether the transition probabilities are known or not. We also provide a gradient-based policy improvement algorithm that converges to a local optimum of the risk objective.
Explore related subjects
Keep this discovery
Pengqian Yu, Jia Yuan Yu, Huan Xu. 2015-12-02. Central-limit approach to risk-aware Markov decision processes. https://arxiv.org/abs/1512.00583
Cite the original work for its findings. Save a collection to share your selection of sources.