arXiv · 2402.18836
A Model-Based Approach for Improving Reinforcement Learning Efficiency Leveraging Expert Observations
Abstract
This paper investigates how to incorporate expert observations (without explicit information on expert actions) into a deep reinforcement learning setting to improve sample efficiency. First, we formulate an augmented policy loss combining a maximum entropy reinforcement learning objective with a behavioral cloning loss that leverages a forward dynamics model. Then, we propose an algorithm that automatically adjusts the weights of each component in the augmented loss function. Experiments on a variety of continuous control tasks demonstrate that the proposed algorithm outperforms various benchmarks by effectively utilizing available expert observations.
Explore related subjects
Keep this discovery
Erhan Can Ozcan, Vittorio Giammarino, James Queeney, Ioannis Ch. Paschalidis. 2024-02-29. A Model-Based Approach for Improving Reinforcement Learning Efficiency Leveraging Expert Observations. https://doi.org/10.1109/cdc56724.2024.10885921
Cite the original work for its findings. Save a collection to share your selection of sources.