TY - RPRT TI - Is Vanilla Policy Gradient Overlooked? Analyzing Deep Reinforcement Learning for Hanabi AU - Bram Grooten AU - Jelle Wemmenhove AU - Maurice Poot AU - Jim Portegies PY - 2022 UR - https://arxiv.org/abs/2203.11656 ID - 2203.11656 ER -