TY - RPRT TI - Understanding Gap-Dependent Regret for Optimism-Based Reinforcement Learning with Linear Function Approximation AU - Haochen Zhang AU - Zhong Zheng AU - Lingzhou Xue PY - 2026 UR - https://arxiv.org/abs/2602.20297 ID - 2602.20297 ER -