TY - RPRT TI - Logarithmic Regret for Reinforcement Learning with Linear Function Approximation AU - Jiafan He AU - Dongruo Zhou AU - Quanquan Gu PY - 2021 UR - https://arxiv.org/abs/2011.11566 ID - 2011.11566 ER -