arXiv · 2405.20538
Q-learning as a monotone scheme
Abstract
Stability issues with reinforcement learning methods persist. To better understand some of these stability and convergence issues involving deep reinforcement learning methods, we examine a simple linear quadratic example. We interpret the convergence criterion of exact Q-learning in the sense of a monotone scheme and discuss consequences of function approximation on monotonicity properties.
Explore related subjects
Keep this discovery
Lingyi Yang. 2024-05-30. Q-learning as a monotone scheme. https://arxiv.org/abs/2405.20538
Cite the original work for its findings. Save a collection to share your selection of sources.