TY - RPRT TI - Safe Exploration of State and Action Spaces in Reinforcement Learning AU - Javier Garcia AU - Fernando Fernandez PY - 2014 DO - 10.1613/jair.3761 UR - https://arxiv.org/abs/1402.0560 ID - 1402.0560 ER -