TY - RPRT TI - Provably Good Batch Reinforcement Learning Without Great Exploration AU - Yao Liu AU - Adith Swaminathan AU - Alekh Agarwal AU - Emma Brunskill PY - 2020 UR - https://arxiv.org/abs/2007.08202 ID - 2007.08202 ER -