TY - RPRT TI - Self-Supervised Online Reward Shaping in Sparse-Reward Environments AU - Farzan Memarian AU - Wonjoon Goo AU - Rudolf Lioutikov AU - Scott Niekum AU - Ufuk Topcu PY - 2021 UR - https://arxiv.org/abs/2103.04529 ID - 2103.04529 ER -