TY - RPRT TI - A Generalized Projected Bellman Error for Off-policy Value Estimation in Reinforcement Learning AU - Andrew Patterson AU - Adam White AU - Martha White PY - 2024 UR - https://arxiv.org/abs/2104.13844 ID - 2104.13844 ER -