TY - RPRT TI - Settling the Horizon-Dependence of Sample Complexity in Reinforcement Learning AU - Yuanzhi Li AU - Ruosong Wang AU - Lin F. Yang PY - 2021 UR - https://arxiv.org/abs/2111.00633 ID - 2111.00633 ER -