TY - RPRT TI - Two-step reinforcement learning for model-free redesign of nonlinear optimal regulator AU - Mei Minami AU - Yuka Masumoto AU - Yoshihiro Okawa AU - Tomotake Sasaki AU - Yutaka Hori PY - 2023 DO - 10.1080/18824889.2023.2278753 UR - https://arxiv.org/abs/2103.03808 ID - 2103.03808 ER -