TY - RPRT TI - Shaping Rewards for Reinforcement Learning with Imperfect Demonstrations using Generative Models AU - Yuchen Wu AU - Melissa Mozifian AU - Florian Shkurti PY - 2020 UR - https://arxiv.org/abs/2011.01298 ID - 2011.01298 ER -