TY - RPRT TI - Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithms AU - Miao Lu AU - Han Zhong AU - Tong Zhang AU - Jose Blanchet PY - 2026 UR - https://arxiv.org/abs/2404.03578 ID - 2404.03578 ER -