TY - RPRT TI - MUSBO: Model-based Uncertainty Regularized and Sample Efficient Batch Optimization for Deployment Constrained Reinforcement Learning AU - DiJia Su AU - Jason D. Lee AU - John M. Mulvey AU - H. Vincent Poor PY - 2021 UR - https://arxiv.org/abs/2102.11448 ID - 2102.11448 ER -