TY - RPRT TI - Offline Meta-Reinforcement Learning with Advantage Weighting AU - Eric Mitchell AU - Rafael Rafailov AU - Xue Bin Peng AU - Sergey Levine AU - Chelsea Finn PY - 2021 UR - https://arxiv.org/abs/2008.06043 ID - 2008.06043 ER -