TY - RPRT TI - Fast Policy Learning through Imitation and Reinforcement AU - Ching-An Cheng AU - Xinyan Yan AU - Nolan Wagener AU - Byron Boots PY - 2018 UR - https://arxiv.org/abs/1805.10413 ID - 1805.10413 ER -