TY - RPRT TI - Learning the optimal state-feedback via supervised imitation learning AU - Dharmesh Tailor AU - Dario Izzo PY - 2019 UR - https://arxiv.org/abs/1901.02369 ID - 1901.02369 ER -