TY - RPRT TI - Abstract Demonstrations and Adaptive Exploration for Efficient and Stable Multi-step Sparse Reward Reinforcement Learning AU - Xintong Yang AU - Ze Ji AU - Jing Wu AU - Yu-kun Lai PY - 2022 DO - 10.1109/icac55051.2022.9911100 UR - https://arxiv.org/abs/2207.09243 ID - 2207.09243 ER -