TY - RPRT TI - Dual Self-Awareness Value Decomposition Framework without Individual Global Max for Cooperative Multi-Agent Reinforcement Learning AU - Zhiwei Xu AU - Bin Zhang AU - Dapeng Li AU - Guangchong Zhou AU - Zeren Zhang AU - Guoliang Fan PY - 2023 UR - https://arxiv.org/abs/2302.02180 ID - 2302.02180 ER -