TY - RPRT TI - Constrained Style Learning from Imperfect Demonstrations under Task Optimality AU - Kehan Wen AU - Chenhao Li AU - Junzhe He AU - Marco Hutter PY - 2025 UR - https://arxiv.org/abs/2507.09371 ID - 2507.09371 ER -