TY - RPRT TI - Continuous-time reinforcement learning: ellipticity enables model-free value function approximation AU - Wenlong Mou PY - 2026 UR - https://arxiv.org/abs/2602.06930 ID - 2602.06930 ER -