arXiv · 1802.07668
A model for system uncertainty in reinforcement learning
Abstract
This work provides a rigorous framework for studying continuous time control problems in uncertain environments. The framework considered models uncertainty in state dynamics as a measure on the space of functions. This measure is considered to change over time as agents learn their environment. This model can be seem as a variant of either Bayesian reinforcement learning or adaptive control. We study necessary conditions for locally optimal trajectories within this model, in particular deriving an appropriate dynamic programming principle and Hamilton-Jacobi equations. This model provides one possible framework for studying the tradeoff between exploration and exploitation in reinforcement learning.
Explore related subjects
Keep this discovery
Ryan Murray, Michele Palladino. 2018-02-21. A model for system uncertainty in reinforcement learning. https://arxiv.org/abs/1802.07668
Cite the original work for its findings. Save a collection to share your selection of sources.