TY - RPRT TI - On Function Approximation in Reinforcement Learning: Optimism in the Face of Large State Spaces AU - Zhuoran Yang AU - Chi Jin AU - Zhaoran Wang AU - Mengdi Wang AU - Michael I. Jordan PY - 2020 UR - https://arxiv.org/abs/2011.04622 ID - 2011.04622 ER -