TY - RPRT TI - Approximation Benefits of Policy Gradient Methods with Aggregated States AU - Daniel Russo PY - 2022 UR - https://arxiv.org/abs/2007.11684 ID - 2007.11684 ER -