TY - RPRT TI - Opponent Aware Reinforcement Learning AU - Victor Gallego AU - Roi Naveiro AU - David Rios Insua AU - David Gomez-Ullate Oteiza PY - 2026 DO - 10.1016/j.ejor.2026.08.031 UR - https://arxiv.org/abs/1908.08773 ID - 1908.08773 ER -