TY - RPRT TI - On the continuity and smoothness of the value function in reinforcement learning and optimal control AU - Hans Harder AU - Sebastian Peitz PY - 2024 UR - https://arxiv.org/abs/2403.14432 ID - 2403.14432 ER -