arXiv · 2002.12446
Provably Efficient Third-Person Imitation from Offline Observation
Abstract
Domain adaptation in imitation learning represents an essential step towards improving generalizability. However, even in the restricted setting of third-person imitation where transfer is between isomorphic Markov Decision Processes, there are no strong guarantees on the performance of transferred policies. We present problem-dependent, statistical learning guarantees for third-person imitation from observation in an offline setting, and a lower bound on performance in the online setting.
Explore related subjects
Keep this discovery
Aaron Zweig, Joan Bruna. 2020-02-27. Provably Efficient Third-Person Imitation from Offline Observation. https://arxiv.org/abs/2002.12446
Cite the original work for its findings. Save a collection to share your selection of sources.