TY - RPRT TI - Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning AU - Zijun Chen AU - Shengbo Wang AU - Nian Si PY - 2026 UR - https://arxiv.org/abs/2505.10007 ID - 2505.10007 ER -