TY - RPRT TI - Accuracy of Discretely Sampled Stochastic Policies in Continuous-time Reinforcement Learning AU - Yanwei Jia AU - Du Ouyang AU - Yufei Zhang PY - 2025 UR - https://arxiv.org/abs/2503.09981 ID - 2503.09981 ER -