TY - RPRT TI - Statistically Efficient Advantage Learning for Offline Reinforcement Learning in Infinite Horizons AU - Chengchun Shi AU - Shikai Luo AU - Yuan Le AU - Hongtu Zhu AU - Rui Song PY - 2022 UR - https://arxiv.org/abs/2202.13163 ID - 2202.13163 ER -