TY - RPRT TI - STIR$^2$: Reward Relabelling for combined Reinforcement and Imitation Learning on sparse-reward tasks AU - Jesus Bujalance Martin AU - Fabien Moutarde PY - 2023 UR - https://arxiv.org/abs/2201.03834 ID - 2201.03834 ER -