TY - RPRT TI - Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables AU - Kate Rakelly AU - Aurick Zhou AU - Deirdre Quillen AU - Chelsea Finn AU - Sergey Levine PY - 2019 UR - https://arxiv.org/abs/1903.08254 ID - 1903.08254 ER -