TY - RPRT TI - End-to-End Multi-Task Policy Learning from NMPC for Quadruped Locomotion AU - Anudeep Sajja AU - Shahram Khorshidi AU - Sebastian Houben AU - Maren Bennewitz PY - 2025 UR - https://arxiv.org/abs/2505.08574 ID - 2505.08574 ER -