TY - RPRT TI - Safeguarded Progress in Reinforcement Learning: Safe Bayesian Exploration for Control Policy Synthesis AU - Rohan Mitta AU - Hosein Hasanbeig AU - Jun Wang AU - Daniel Kroening AU - Yiannis Kantaros AU - Alessandro Abate PY - 2023 UR - https://arxiv.org/abs/2312.11314 ID - 2312.11314 ER -