TY - RPRT TI - Deviations from the Nash equilibrium in a two-player optimal execution game with reinforcement learning AU - Fabrizio Lillo AU - Andrea Macrì PY - 2026 UR - https://arxiv.org/abs/2408.11773 ID - 2408.11773 ER -