TY - RPRT TI - Reinforcement Learning in Linear Quadratic Deep Structured Teams: Global Convergence of Policy Gradient Methods AU - Vida Fathi AU - Jalal Arabneydi AU - Amir G. Aghdam PY - 2020 UR - https://arxiv.org/abs/2011.14393 ID - 2011.14393 ER -