arXiv · 1303.1942
Optimal Control of MDPs with Temporal Logic Constraints
Abstract
In this paper, we focus on formal synthesis of control policies for finite Markov decision processes with non-negative real-valued costs. We develop an algorithm to automatically generate a policy that guarantees the satisfaction of a correctness specification expressed as a formula of Linear Temporal Logic, while at the same time minimizing the expected average cost between two consecutive satisfactions of a desired property. The existing solutions to this problem are sub-optimal. By leveraging ideas from automata-based model checking and game theory, we provide an optimal solution. We demonstrate the approach on an illustrative example.
Explore related subjects
Keep this discovery
Maria Svorenova, Ivana Cerna, Calin Belta. 2013-09-09. Optimal Control of MDPs with Temporal Logic Constraints. https://arxiv.org/abs/1303.1942
Cite the original work for its findings. Save a collection to share your selection of sources.