TY - RPRT TI - The intrinsic motivation of reinforcement and imitation learning for sequential tasks AU - Sao Mai Nguyen PY - 2024 UR - https://arxiv.org/abs/2412.20573 ID - 2412.20573 ER -