TY - RPRT TI - An Output Feedback Q-learning Algorithm for Optimal Control of Nonlinear Systems with Koopman Linear Embedding AU - Victor G. Lopez AU - Malte Heinrich AU - Matthias A. Müller PY - 2026 UR - https://arxiv.org/abs/2603.29858 ID - 2603.29858 ER -