arXiv · 2502.13697
Linear programming for finite-horizon vector-valued Markov decision processes
Abstract
We propose a vector linear programming formulation for a non-stationary, finite-horizon Markov decision process with vector-valued rewards. Pareto efficient policies are shown to correspond to efficient solutions of the linear program, and vector linear programming theory allows us to fully characterize deterministic efficient policies. An algorithm for enumerating all efficient deterministic policies is presented then tested numerically in an engineering application.
Explore related subjects
Keep this discovery
Anas Mifrani, Dominikus Noll. 2025-02-19. Linear programming for finite-horizon vector-valued Markov decision processes. https://doi.org/10.61208/pjo-2025-028
Cite the original work for its findings. Save a collection to share your selection of sources.