TY - RPRT TI - F2A2: Flexible Fully-decentralized Approximate Actor-critic for Cooperative Multi-agent Reinforcement Learning AU - Wenhao Li AU - Bo Jin AU - Xiangfeng Wang AU - Junchi Yan AU - Hongyuan Zha PY - 2023 UR - https://arxiv.org/abs/2004.11145 ID - 2004.11145 ER -