TY - RPRT TI - Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees AU - Nan Jiang AU - Tengyang Xie PY - 2025 UR - https://arxiv.org/abs/2510.04088 ID - 2510.04088 ER -