TY - RPRT TI - Solving Sokoban with forward-backward reinforcement learning AU - Yaron Shoham AU - Gal Elidan PY - 2021 UR - https://arxiv.org/abs/2105.01904 ID - 2105.01904 ER -