TY - RPRT TI - Provably adaptive reinforcement learning in metric spaces AU - Tongyi Cao AU - Akshay Krishnamurthy PY - 2021 UR - https://arxiv.org/abs/2006.10875 ID - 2006.10875 ER -