arXiv · 1401.3907
Policy Invariance under Reward Transformations for General-Sum Stochastic Games
Abstract
We extend the potential-based shaping method from Markov decision processes to multi-player general-sum stochastic games. We prove that the Nash equilibria in a stochastic game remains unchanged after potential-based shaping is applied to the environment. The property of policy invariance provides a possible way of speeding convergence when learning to play a stochastic game.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Xiaosong Lu, Howard M. Schwartz, Sidney N. Givigi Jr. 2014-01-16. Policy Invariance under Reward Transformations for General-Sum Stochastic Games. https://doi.org/10.1613/jair.3384
Cite the original work for its findings. Save a collection to share your selection of sources.