arXiv · 1607.00678
Optimizing the Expected Mean Payoff in Energy Markov Decision Processes
Abstract
Energy Markov Decision Processes (EMDPs) are finite-state Markov decision processes where each transition is assigned an integer counter update and a rational payoff. An EMDP configuration is a pair s(n), where s is a control state and n is the current counter value. The configurations are changed by performing transitions in the standard way. We consider the problem of computing a safe strategy (i.e., a strategy that keeps the counter non-negative) which maximizes the expected mean payoff.
Explore related subjects
Keep this discovery
Tomáš Brázdil, Antonín Kučera, Petr Novotný. 2016-07-03. Optimizing the Expected Mean Payoff in Energy Markov Decision Processes. https://arxiv.org/abs/1607.00678
Cite the original work for its findings. Save a collection to share your selection of sources.