TY - RPRT TI - Sample-Efficient Online Control Policy Learning with Real-Time Recursive Model Updates AU - Zixin Zhang AU - James Avtges AU - Todd D. Murphey PY - 2025 UR - https://arxiv.org/abs/2509.08241 ID - 2509.08241 ER -