arXiv · 2507.06582
Learning controllable dynamics through informative exploration
Abstract
Environments with controllable dynamics are usually understood in terms of explicit models. However, such models are not always available, but may sometimes be learned by exploring an environment. In this work, we investigate using an information measure called "predicted information gain" to determine the most informative regions of an environment to explore next. Applying methods from reinforcement learning allows good suboptimal exploring policies to be found, and leads to reliable estimates of the underlying controllable dynamics. This approach is demonstrated by comparing with several myopic exploration approaches.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Peter N. Loxley, Friedrich T. Sommer. 2025-07-09. Learning controllable dynamics through informative exploration. https://arxiv.org/abs/2507.06582
Cite the original work for its findings. Save a collection to share your selection of sources.