TY - RPRT TI - Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning AU - Huizhen Yu AU - Yi Wan AU - Richard S. Sutton PY - 2025 DO - 10.1137/25m1769806 UR - https://arxiv.org/abs/2409.03915 ID - 2409.03915 ER -