TY - RPRT TI - TacticZero: Learning to Prove Theorems from Scratch with Deep Reinforcement Learning AU - Minchao Wu AU - Michael Norrish AU - Christian Walder AU - Amir Dezfouli PY - 2021 UR - https://arxiv.org/abs/2102.09756 ID - 2102.09756 ER -