TY - RPRT TI - Efficient Off-Policy Q-Learning for Data-Based Discrete-Time LQR Problems AU - Victor G. Lopez AU - Mohammad Alsalti AU - Matthias A. Müller PY - 2023 DO - 10.1109/tac.2023.3235967 UR - https://arxiv.org/abs/2105.07761 ID - 2105.07761 ER -