TY - RPRT TI - The RL Perceptron: Generalisation Dynamics of Policy Learning in High Dimensions AU - Nishil Patel AU - Sebastian Lee AU - Stefano Sarao Mannelli AU - Sebastian Goldt AU - Andrew Saxe PY - 2024 DO - 10.1103/physrevx.15.021051 UR - https://arxiv.org/abs/2306.10404 ID - 2306.10404 ER -