arXiv · 1703.07075
Pseudorehearsal in value function approximation
Abstract
Catastrophic forgetting is of special importance in reinforcement learning, as the data distribution is generally non-stationary over time. We study and compare several pseudorehearsal approaches for Q-learning with function approximation in a pole balancing task. We have found that pseudorehearsal seems to assist learning even in such very simple problems, given proper initialization of the rehearsal parameters.
Explore related subjects
Keep this discovery
Vladimir Marochko, Leonard Johard, Manuel Mazzara. 2017-03-21. Pseudorehearsal in value function approximation. https://arxiv.org/abs/1703.07075
Cite the original work for its findings. Save a collection to share your selection of sources.