TY - RPRT TI - Distance-rank Aware Sequential Reward Learning for Inverse Reinforcement Learning with Sub-optimal Demonstrations AU - Lu Li AU - Yuxin Pan AU - Ruobing Chen AU - Jie Liu AU - Zilin Wang AU - Yu Liu AU - Zhiheng Li PY - 2023 UR - https://arxiv.org/abs/2310.08823 ID - 2310.08823 ER -