TY - RPRT TI - Approximate exploitability: Learning a best response in large games AU - Finbarr Timbers AU - Nolan Bard AU - Edward Lockhart AU - Marc Lanctot AU - Martin Schmid AU - Neil Burch AU - Julian Schrittwieser AU - Thomas Hubert AU - Michael Bowling PY - 2022 UR - https://arxiv.org/abs/2004.09677 ID - 2004.09677 ER -