TY - RPRT TI - One-Shot High-Fidelity Imitation: Training Large-Scale Deep Nets with RL AU - Tom Le Paine AU - Sergio Gómez Colmenarejo AU - Ziyu Wang AU - Scott Reed AU - Yusuf Aytar AU - Tobias Pfaff AU - Matt W. Hoffman AU - Gabriel Barth-Maron AU - Serkan Cabi AU - David Budden AU - Nando de Freitas PY - 2018 UR - https://arxiv.org/abs/1810.05017 ID - 1810.05017 ER -