TY - RPRT TI - Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach AU - Minghao Han AU - Lixian Zhang AU - Chenliang Liu AU - Zhipeng Zhou AU - Jun Wang AU - Wei Pan PY - 2026 UR - https://arxiv.org/abs/2603.00043 ID - 2603.00043 ER -