TY - RPRT TI - $\sqrt{n}$-Regret for Learning in Markov Decision Processes with Function Approximation and Low Bellman Rank AU - Kefan Dong AU - Jian Peng AU - Yining Wang AU - Yuan Zhou PY - 2020 UR - https://arxiv.org/abs/1909.02506 ID - 1909.02506 ER -