arXiv · 2211.07888
Partially Observable Discrete-time Discounted Markov Games with General Utility
Abstract
In this paper, we investigate a partially observable zero sum games where the state process is a discrete time Markov chain. We consider a general utility function in the optimization criterion. We show the existence of value for both finite and infinite horizon games and also establish the existence of optimal polices. The main step involves converting the partially observable game into a completely observable game which also keeps track of the total discounted accumulated reward/cost.
Explore related subjects
Keep this discovery
Arnab Bhabak, Subhamay saha. 2022-11-15. Partially Observable Discrete-time Discounted Markov Games with General Utility. https://arxiv.org/abs/2211.07888
Cite the original work for its findings. Save a collection to share your selection of sources.