TY - RPRT TI - Decentralized Natural Policy Gradient with Variance Reduction for Collaborative Multi-Agent Reinforcement Learning AU - Jinchi Chen AU - Jie Feng AU - Weiguo Gao AU - Ke Wei PY - 2022 UR - https://arxiv.org/abs/2209.02179 ID - 2209.02179 ER -