TY - RPRT TI - Outcome-Driven Reinforcement Learning via Variational Inference AU - Tim G. J. Rudner AU - Vitchyr H. Pong AU - Rowan McAllister AU - Yarin Gal AU - Sergey Levine PY - 2022 UR - https://arxiv.org/abs/2104.10190 ID - 2104.10190 ER -