TY - RPRT TI - Safe reinforcement learning for probabilistic reachability and safety specifications: A Lyapunov-based approach AU - Subin Huh AU - Insoon Yang PY - 2020 UR - https://arxiv.org/abs/2002.10126 ID - 2002.10126 ER -