arXiv · 2606.01363
All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning
Abstract
Model-based reinforcement learning (MBRL) infers information about the environment from a learned dynamics model and bears the potential to address open problems such as data efficient and safe learning in robotics. However, inaccuracies of the learned dynamics model are typically exploited by the agent, substantially hampering the capabilities of MBRL methods. We present a framework for dealing with inaccuracies of probabilistic models through targeted handling of uncertainty that effectively mitigates model exploitation. We present recent successes in learning directly on hardware and safe exploration, and discuss future directions for uncertainty-aware MBRL.
Explore related subjects
Keep this discovery
Bernd Frauenknecht, Devdutt Subhasish, Artur Eisele, Friedrich Solowjow, Sebastian Trimpe. 2026-05-31. All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning. https://arxiv.org/abs/2606.01363
Cite the original work for its findings. Save a collection to share your selection of sources.