TY - RPRT TI - Near-Minimax-Optimal Distributional Reinforcement Learning with a Generative Model AU - Mark Rowland AU - Li Kevin Wenliang AU - Rémi Munos AU - Clare Lyle AU - Yunhao Tang AU - Will Dabney PY - 2024 UR - https://arxiv.org/abs/2402.07598 ID - 2402.07598 ER -