TY - RPRT TI - Uniform-PAC Bounds for Reinforcement Learning with Linear Function Approximation AU - Jiafan He AU - Dongruo Zhou AU - Quanquan Gu PY - 2021 UR - https://arxiv.org/abs/2106.11612 ID - 2106.11612 ER -