arXiv · 2010.06794
Minimax Q-learning Control for Linear Systems Using the Wasserstein Metric
Abstract
Stochastic optimal control usually requires an explicit dynamical model with probability distributions, which are difficult to obtain in practice. In this work, we consider the linear quadratic regulator (LQR) problem of unknown linear systems and adopt a Wasserstein penalty to address the distribution uncertainty of additive stochastic disturbances. By constructing an equivalent deterministic game of the penalized LQR problem, we propose a Q-learning method with convergence guarantees to learn an optimal minimax controller.
Explore related subjects
Keep this discovery
Feiran Zhao, Keyou You. 2020-10-14. Minimax Q-learning Control for Linear Systems Using the Wasserstein Metric. https://arxiv.org/abs/2010.06794
Cite the original work for its findings. Save a collection to share your selection of sources.