TY - RPRT TI - An efficient data-based off-policy Q-learning algorithm for optimal output feedback control of linear systems AU - Mohammad Alsalti AU - Victor G. Lopez AU - Matthias A. Müller PY - 2024 UR - https://arxiv.org/abs/2312.03451 ID - 2312.03451 ER -