TY - RPRT TI - Robust Reinforcement Learning using Least Squares Policy Iteration with Provable Performance Guarantees AU - Kishan Panaganti AU - Dileep Kalathil PY - 2021 UR - https://arxiv.org/abs/2006.11608 ID - 2006.11608 ER -