TY - RPRT TI - Reinforcement Learning For Constraint Satisfaction Game Agents (15-Puzzle, Minesweeper, 2048, and Sudoku) AU - Anav Mehta PY - 2021 UR - https://arxiv.org/abs/2102.06019 ID - 2102.06019 ER -