TY - RPRT TI - Convergence and Optimality of Policy Gradient Methods in Weakly Smooth Settings AU - Matthew S. Zhang AU - Murat A. Erdogdu AU - Animesh Garg PY - 2022 UR - https://arxiv.org/abs/2111.00185 ID - 2111.00185 ER -